Four New Major Open Source Foundation Models in a Week
Was this email forwarded to you? Sign up here Four New Major Open Source Foundation Models in a WeekDBRX, Grok 1.5, Samba-CoE and Jamba are all bringing unique innovations to open source generative AI.Next Week in The Sequence:
You can subscribe to The Sequence using the link below:📝 Editorial: Four New Major Open Source Foundation Models in a WeekOpen source generative AI is experiencing tremendous momentum, and last week was a major example of this with the release of four major foundation models. By open source, we refer to the weights of the models and not the training datasets or processes. At this time, it's fair to say that the model weights are where most companies draw the line between open source and closed source. Many purists do not consider this true open source, but in a field evolving as rapidly as generative AI, preserving a level of competitive advantage is essential for any company. Let’s just say that the nature of open source is being reimagined for generative AI. The fast pace of generative AI also makes the open source race even more fascinating. Last week, we witnessed the release of four major open source models, each innovative in its own way:
Regardless of where you fall in the commercial vs. open source debate in generative AI, it is undeniable that the latter will play a major role in the mainstream adoption of this technology. This week shows how strong the momentum in open source generative AI is." With just one week left until apply() ‘24, the premier virtual conference for engineers mastering AI and ML, we wanted to remind you to secure your spot before it's too late! Date: Wednesday, April 3 / 9:00AM – 5:00PM PT / Virtual At apply(), our goal is to provide you with the tools and insights you need to conquer AI and ML challenges at production scale. With speakers from LangChain, Meta, Pinterest, Vanguard, Visa, Samsung, NextDoor, and many more in the lineup, this year's event promises to be our best yet. Be sure to join live for the chance to win swag or a giveaway prize! 🔎 ML ResearchCan LLMs Explore?Researchers from Microsoft and Carnegie Mellon University published a paper exploring the intriguing thesis of LLM’s ability to engage in exploration, an ability typically reserved for reinforcement learning models. The research describes environments such as multi-armed bandits in prompts and determine whether LLMs can explore the environment in order to take actions —> Read more. Tnt-LLMMicrosoft Research published a paper introducing Tnt-LLM, an LLM framework that generates and predict task labels with minimum user involvement. Tnt-LLM is actively used to discover Microsoft CoPilot’s user’s intent —> Read more. AutoBNNGoogle Research published and research adn open sourced AutoBNN, a JAX framework for interpretable time series forecasting models. AutoBNN’s core idea is to combine the interpretability of traditional time series models with the scalability of neural networks in a single architecture —> Read more. SaLEMAmazon Science published a paper introducing SaLEM (for salient-layers editing model), a method for editing layers in an LLM. SaLEM’s key contribution is that it can actually select the layers to be edited automatically —> Read more. SAFEGoogle DeepMind published a paper presenting Search-Augmented Factuality Evaluator (SAFE), a method for factual evaluation in LLMs using synthetic data. SAFE breaks down a long LLM response into specific facts and evaluates its individual accuracy —> Read more. 🤖 Cool AI Tech ReleasesDBRXDatabricks released DBRX, a new state-of-the-art open source LLM —> Read more. JambaAI21 Labs open sourced Jamba, a new model that augments Structured State Space model (SSM) with elements of the transformer architecture —> Read more. Samba CoE v0.2Samba Nova previewed the performance of Samba CoE v0.2, a new version of Samba-1 which scored incredibly high across many benchmarks —> Read more. Grok 1.5X.ai released Grok 1.5 with improved content reasoning capabilities and larger content length —> Read more. Voice EngineOpenAI published some details about Voice Engine, a new model for creating custom voices —> Read more. 🛠 Real World MLVideo Content Moderation at YelpYelp discusses the ML architecture powering its video content moderation solution —> Read more. 📡AI Radar
You’re on the free list for TheSequence Scope and TheSequence Chat. For the full experience, become a paying subscriber to TheSequence Edge. Trusted by thousands of subscribers from the leading AI labs and universities. |
Older messages
Edge 381: Google DeepMind's PrompBreeder Self-Improves Prompts
Thursday, March 28, 2024
The method combines chain of thoughts, plan and solve and evolutionary algorithms in a single mthod. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
Edge 380: A New Series About Autonomous Agents
Tuesday, March 26, 2024
The series will cover memory, action execution, planning, collaboration and many other characteristics of autonomous agents. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
📝 Guest Post: Zilliz Unveiled Milvus 2.4 at GTC 24, Transforming Vector Databases with GPU Acceleration*
Monday, March 25, 2024
Collaboration with NVIDIA boosts Milvus performance 50x ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
NVIDIA’s GTC in Four Headlines
Sunday, March 24, 2024
Impressive AI hardware innovations and interesting software moves. ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏ ͏
📌 Exciting lineup for apply() 2024 is now live
Friday, March 22, 2024
Exciting news! The agenda for apply() 2024, Tecton's premier virtual conference dedicated to mastering AI and ML at production scale, is now live! Join us on Wednesday, April 3, for a day packed
You Might Also Like
📧 Building Async APIs in ASP.NET Core - The Right Way
Saturday, November 23, 2024
Building Async APIs in ASP .NET Core - The Right Way Read on: my website / Read time: 5 minutes The .NET Weekly is brought to you by: Even the smartest AI in the world won't save you from a
WebAIM November 2024 Newsletter
Friday, November 22, 2024
WebAIM November 2024 Newsletter Read this newsletter online at https://webaim.org/newsletter/2024/november Features Using Severity Ratings to Prioritize Web Accessibility Remediation When it comes to
➡️ Why Your Phone Doesn't Want You to Sideload Apps — Setting the Default Gateway in Linux
Friday, November 22, 2024
Also: Hey Apple, It's Time to Upgrade the Macs Storage, and More! How-To Geek Logo November 22, 2024 Did You Know Fantasy author JRR Tolkien is credited with inventing the main concept of orcs and
JSK Daily for Nov 22, 2024
Friday, November 22, 2024
JSK Daily for Nov 22, 2024 View this email in your browser A community curated daily e-mail of JavaScript news React E-Commerce App for Digital Products: Part 4 (Creating the Home Page) This component
Spyglass Dispatch: The Fate of Chrome • Amazon Tops Up Anthropic • Pros Quit Xitter • Brave Powers AI Search • Apple's Lazy AI River • RIP Enrique Allen
Friday, November 22, 2024
The Fate of Chrome • Amazon Tops Up Anthropic • Pros Quit Xitter • Brave Powers AI Search • Apple's Lazy AI River • RIP Enrique Allen The Spyglass Dispatch is a free newsletter sent out daily on
Charted | How the Global Distribution of Wealth Has Changed (2000-2023) 💰
Friday, November 22, 2024
This graphic illustrates the shifts in global wealth distribution between 2000 and 2023. View Online | Subscribe | Download Our App Presented by: MSCI >> Get the Free Investor Guide Now FEATURED
Daily Coding Problem: Problem #1616 [Easy]
Friday, November 22, 2024
Daily Coding Problem Good morning! Here's your coding interview problem for today. This problem was asked by Alibaba. Given an even number (greater than 2), return two prime numbers whose sum will
The problem to solve
Friday, November 22, 2024
Use problem framing to define the problem to solve This week, Tom Parson and Krishna Raha share tools and frameworks to identify and address challenges effectively, while Voltage Control highlights
Issue #568: Random mazes, train clock, and ReKill
Friday, November 22, 2024
View this email in your browser Issue #568 - November 22nd 2024 Weekly newsletter about Web Game Development. If you have anything you want to share with our community please let me know by replying to
Whats Next for AI: Interpreting Anthropic CEOs Vision
Friday, November 22, 2024
Top Tech Content sent at Noon! How the world collects web data Read this email in your browser How are you, @newsletterest1? 🪐 What's happening in tech today, November 22, 2024? The HackerNoon