Robotics is Inching Towards it ChatGPT Moment
Was this email forwarded to you? Sign up here Robotics is Inching Towards it ChatGPT MomentMajor developments in robotics from NVIDIA, Meta and MIT.Next Week in The Sequence:
You can subscribe to The Sequence below:📝 Editorial: Robotics is Inching Towards it ChatGPT MomentThe field of AI robotics is currently experiencing a surge in innovation, with researchers developing new techniques and technologies that are pushing the boundaries of what robots can do. One of the most exciting areas of development is the use of large language models (LLMs) to train robots. LLMs are a type of AI that are trained on massive datasets of text and code, and they have shown remarkable ability to generate text, translate languages, and write different kinds of creative content. Researchers are now exploring how to use LLMs to train robots to perform a wide range of tasks, from simple household chores to complex industrial operations. This week we saw several major research contributions in the field of robotics from NVIDIA, MIT and Meta among others. A major challenge in robotics is the heterogeneity of data. Robots generate data from a variety of sources, including vision sensors, robotic arm position encoders, and simulations. These data are often difficult to combine and use to train robots. Researchers at MIT have developed a new technique called Heterogeneous Pretrained Transformers (HPT) that addresses this challenge. HPT aligns data from different sources into a shared "language" that a generative AI model can process. This approach allows robots to be trained on a much larger and more diverse dataset, which can lead to significant improvements in performance. Beyond the technical advancements, the industry is witnessing a growing focus on the integration of touch perception, dexterity, and human-robot interaction. Meta's Fundamental AI Research (FAIR) team is actively working on creating embodied AI agents capable of perceiving and interacting with their surroundings, while also coexisting safely with humans. Their efforts are leading to advancements in areas such as tactile sensing, which allows robots to "feel" and manipulate objects with greater precision. This is exemplified by their development of Meta Sparsh, a general-purpose touch representation that works across various sensors and tasks, and Meta Digit 360, a breakthrough tactile fingertip with human-level multimodal sensing capabilities. The drive towards more versatile and adaptable robots is also evident in the development of new control frameworks. NVIDIA's research on HOVER (Humanoid Versatile Controller) showcases a multi-mode policy distillation framework that consolidates diverse control modes into a unified policy. HOVER allows robots to seamlessly switch between different control modes, such as navigation, manipulation, and human interaction, without the need for retraining.8 This development marks a significant step toward creating more flexible and adaptable robots that can perform a wide range of tasks. The advancements in AI robotics, as highlighted by these recent developments, demonstrate a clear momentum in the field. With the continuous development of new techniques and technologies, we can expect even more impressive progress in the near future. These breakthroughs not only promise to revolutionize industries but also hold the potential to significantly enhance our daily lives. 📍 EventYou’re invited to an exclusive fireside chat with Ben Orkin, VP of Engineering - MLOps at North, hosted by Tecton and Data Science Connect. Discover how this leading fintech company leveraged Tecton to build a system that detects fraud at scale with millisecond-level response times while adapting to emerging fraud patterns. You’ll learn:
Don't miss this deep dive into building mission-critical ML systems that balance speed, scale, and adaptability! –>Register here. 🔎 ML ResearchHOVERNVIDIA, Carnegie Mellon University, UC Berkeley and other AI research labs published the research around HOVER(Humanoid Versatile Controller), a 1.5 million parameter neural network to control humanoid robots. HOVER is based on a distillation method that extracts various control modes under the same policy —> Read more. NotebookLM AudioGoogle DeepMind published some details about the speech generation technologies behind NotebookLM and Illuminate. The solution included audio generation models such as AudioLM or SoundStream as well as specialized transformers for handling audio tokens —> Read more. Advancing Embodied AIMeta FAIR published several papers and research artifacts advancing different areas of embodied AI. The research includes areas such as perception, dexterity, and human-robot interaction —> Read more. Stealing User Prompts from MoEsGoogle DeepMind published a paper proposing an attack against MoE models that can unveil the user’s input prompt. The core of the technique centers on manipulating the expert routing system within the MoE model to capture the entire input —> Read more. LLMs as Data ScientistsSnowflake AI Research published a paper proposing FeatEng, a benchmark designed to evaluate LLMs in data science tasks such as feature engineering code. The benchmark presents a model with a dataset and a series of prompts and scores the generated code —> Read more. Memorization in LLMsResearchers from Princeton University, Google, Allen AI and University of Illinois published a paper proposing a quantitative approach to measure memorization in LLMs. The paper proposes a bechmark based on Knights and Knaves (K&K) puzzles to evaluate memorization in reasoning tasks —> Read more. 🤖 AI Tech ReleasesChatGPT SearchOpenAI unveiled ChatGPT Search allowing it to search web sources —> Read more. MobileLLMMeta AI open sourced MobileLLM, a foundation model optimized for on-device scenarios —> Read more. TensorFlow 2.18The new version of TensorFlow is out —> Read more. SmolLM2HuggingFace open sourced a series of small models optimized for edge computing —> Read more. 🛠 Real World AIConversational AI at AirbnbAirbnb revealed some details about the architecture powering its conversational AI experiences —> Read more. 📡AI Radar
You’re on the free list for TheSequence Scope and TheSequence Chat. For the full experience, become a paying subscriber to TheSequence Edge. Trusted by thousands of subscribers from the leading AI labs and universities. |
Older messages
📽 Fully Virtual: Agents in Production
Friday, November 1, 2024
Must-see event! ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
Edge 444: Learn About Movie Gen: Meta AI's Amazing Audio-Video Generation Model
Thursday, October 31, 2024
The new model represents an important milestone open source video and audio generation. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
The Sequence Chat: Thinking About Transformers as Computers
Wednesday, October 30, 2024
A different way to reflect about the capabilities of transformers. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
Edge 443: EVERYTHING you Need to Know About State Space Models
Tuesday, October 29, 2024
A summary of our series about the most viable alternative to transformers. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
Anthropic, WOW
Sunday, October 27, 2024
New models, an agent that can interact with your computer and a new code generation tool. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
You Might Also Like
SRE Weekly Issue #456
Monday, December 23, 2024
View on sreweekly.com A message from our sponsor, FireHydrant: On-call during the holidays? Spend more time taking in some R&R and less getting paged. Let alerts make their rounds fairly with our
The Power of an Annual Review & Grammarly acquires Coda
Sunday, December 22, 2024
I am looking for my next role, Zen Browser got a fresh new look, Flipboard introduces Surf, Campsite shuts down, and a lot more in this week's issue of Creativerly. Creativerly The Power of an
Daily Coding Problem: Problem #1645 [Hard]
Sunday, December 22, 2024
Daily Coding Problem Good morning! Here's your coding interview problem for today. This problem was asked by Facebook. Implement regular expression matching with the following special characters: .
PD#606 How concurrecy works: A visual guide
Sunday, December 22, 2024
A programmer had a problem. "I'll solve it with threads!". has Now problems. two he ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
RD#486 (React) Things I Regret Not Knowing Earlier
Sunday, December 22, 2024
Keep coding, stay curious, and remember—you've got this
🎶 GIFs Are Neat, but I Want Clips With Sound — Your Own Linux Desktop in the Cloud
Sunday, December 22, 2024
Also: 9 Games That Were Truly Ahead of Their Time, and More! How-To Geek Logo December 22, 2024 Did You Know Dextrose is another name for glucose, so if you see it listed prominently on the ingredients
o3—the new state-of-the-art reasoning model - Sync #498
Sunday, December 22, 2024
Plus: Nvidia's new tiny AI supercomputer; Veo 2 and Imagen 3; Google and Microsoft release reasoning models; Waymo to begin testing in Tokyo; Apptronik partners with DeepMind; and more! ͏ ͏ ͏ ͏ ͏ ͏
Sunday Digest | Featuring 'The World’s 20 Largest Economies, by GDP (PPP)' 📊
Sunday, December 22, 2024
Every visualization published this week, in one place. Dec 22, 2024 | View Online | Subscribe | VC+ | Download Our App Hello, welcome to your Sunday Digest. This week, we visualized public debt by
Android Weekly #654 🤖
Sunday, December 22, 2024
View in web browser 654 December 22nd, 2024 Articles & Tutorials Sponsored Solving ANRs with OpenTelemetry While OpenTelemetry is the new observability standard, it lacks official support for many
😸 Our interview with Amjad Masad
Sunday, December 22, 2024
Welcome back, builders Product Hunt Sunday, Dec 22 The Roundup This newsletter was brought to you by AssemblyAI Welcome back, builders Happy Sunday! We've got a special edition of the Roundup this