perhaps an unpopular opinion: but if you only think of people as replaceable cogs, then you shouldn't be surprised when those are the only kinds of folks your company can attract
This is one of the clearest satellite views you will ever see of California.
Breathtaking satellite image today of California, one of the most beautiful places on the planet.
As the father of two daughters who participated on @FIRSTweets teams in high school (the older one continued her excitement around robotics & just finished her PhD in robotics and ML a couple of months ago), I'm very happy to see @Googleorg supporting robotics participation!
NVIDIA finally released Neuralangelo's source code!
The model can turn videos from any device into detailed 3D structures, fully replicating buildings, sculptures, or other real aworld objects or spaces virtually.
Here's how it works:
A model utilizes a 2D video with multiple angles of an object or scene. I selects frames from different viewpoints to understand depth, size, and shape.
The AI creates an initial 3D representation, similar to a sculptor shaping a subject. The render is optimized to enhance details, like a sculptor refining texture.
The outcome is a 3D object or scene suitable for virtual reality, digital twins, or robotics.
The famed Stanford Smallville is officially open-source!
25 AI agents inhabit a digital Westworld, unaware that they are living in a simulation. They go to work, gossip, organize socials, make new friends, and even fall in love. Each has unique personality and backstory.
Smallville is among the most inspiring AI agent experiments in 2023. We often talk about a single LLM's emergent abilities, but multi-agent emergence could be way more complex and fascinating at scale. A population of AI can play out the evolution of an entire civilization.
Endless new possibilities ahead. Gaming will be the first to feel the impact.
Github: https://t.co/xUll7KaaTp
Paper: https://t.co/PMDQysrOz9
Authors: @joon_s_pk@joseph_c_obrien@carriejcai@merrierm@percyliang@msbernst
🚀 We've just dropped the first edition of our weekly newsletter summary of the latest #AI & #ML research papers! Curated and made entirely by #GPT4
Subscribe for a free paid week👇
This is a big day. Meta is open-sourcing AudioCraft.
You can now generate incredible music and sounds with a single prompt.
It includes the most performant Generative AI Model (audio) on the market, the "Llama" of Audio.
The research framework contains the weights and code of these models:
▸ MusicGen: controllable text-to-music model.
▸ AudioGen: text-to-sound model.
▸ EnCodec: high fidelity neural audio codec.
▸ Multi Band Diffusion: An EnCodec compatible decoder using diffusion.
This is going to tremendously speed up audio research 👏
LLava just hit 3800 stars on Github.
It's a multimodal Large Language-and-Vision Assistant that can understand images and text.
LLava can even handle memes (the same ones GPT-4 demo'ed at launch) and set a new SOTA on Science QA.
It also supports LLaMA-2, LoRA training with academia GPUs, higher resolution (336x336), 4-/8- inference.
Code Interpreter has been out for 10 days, and it's incredible. It's like having a personal dev capable of running scripts for $20.
Some are even calling it GPT4.5. I think they're right.
Here are the best examples and resources I've found:
Tested Bard converting a screenshot to HTML/CSS.
Here's the input (GoogleStore.jpg, prompt) & output. Still a ways to go with accuracy, but bootstrapping a starting point for developers feels like it will get easier. Excited about the future of multi-modal support.
Google Bard gets Multimodal 🤯
Bard can now extract text summary from the image of an invoice and summarize it in a nice table format.
Witness the OCR magic in action 🔥
Thrilled to share my first proper Milky Way shot in years. This was captured off a small island far from city lights, allowing me to see the incredible detail of our galactic core. If more people could see the night sky like this, I believe the world would be a better place.
one of the easiest policy wins i can imagine for the US is to reform high-skill immigration.
the fact that many of the most talented people in the world want to be here is a hard-won gift; embracing them is the key to keeping it that way.
hard to get this back if we lose it.
There is no greater skill a director can learn than the ability to accurately imagine the possibilities of a location after a talented production designer gets a hold of it.
A quick story:
GPT-Engineer is insane.
Simply specify what you want to build with a prompt, and the AI agent builds the entire codebase.
It's completely free and open-source, and it's already at nearly 17,000 stars on GitHub 🤯
👶🤖 Introducing `gpt-engineer`
▸ One prompt generates a codebase
▸ Asks clarifying questions
▸ Generates technical spec
▸ Writes all necessary code
▸ Easy to add your own reasoning steps, modify, and experiment
▸ open source: https://t.co/NCSp8vsKA0
▸ Lets you finish a coding project in minutes.
@swyx scooped me on releasing first ❤️ – similarities to smol but many new innovations, check README.
Next, make it bootstrap. Which means:
A "gpt engineer" prompt should generate all functionality of the gpt-engineer repo itself.
Horrific to see Uganda pass one of the world’s cruellest anti-LGBTQ+ laws. People could be executed, and thousands exiled just because of who they love. This law is malicious and inhumane https://t.co/xbcOgvDMeT