/
Navigation
Chronicles
Browse all articles
Explore
Semantic exploration
Research
Entity momentum
Nexus
Correlations & relationships
Story Arc
Topic evolution
Drift Map
Semantic trajectory animation
Posts
Analysis & commentary
Pulse API
Tech news intelligence API
Browse
Entities
Companies, people, products, technologies
Domains
Browse by publication source
Handles
Browse by social media handle
Detection
Concept Search
Semantic similarity search
High Impact Stories
Top coverage by position
Sentiment Analysis
Positive/negative coverage
Anomaly Detection
Unusual coverage patterns
Analysis
Rivalry Report
Compare two entities head-to-head
Semantic Pivots
Narrative discontinuities
Crisis Response
Event recovery patterns
Connected
Search: /
Command: ⌘K
Embeddings: large
TEXXR

Chronicles

The story behind the story

days · browse · Enter similar · o open

Mistral announces Mistral Large 2, the new generation of its flagship model, with 123B parameters; commercial usage requires a separate license

Today, we are announcing Mistral Large 2, the new generation of our flagship model. Tobias Mann / The Register : Mistral Large 2 leaps out as a leaner, meaner rival to GPT-4-class AI models MD Ijaj Khan / Hindustan Times : Google integrates Mistral AI's codestral model: What is it and how will it help developers? The Indian Express : Mistral unveils new AI model Large 2, takes on Meta's Llama 3.1 and OpenAI's GPT-4o Luke Jones / WinBuzzer : Mistral Launches 123 Billion Parameter “Mistral Large 2” AI Model Ben's Bites : Mistral adds Large 2 to frontier AI models. Alexey Shabanov / TestingCatalog : Mistral Large 2 now available on major cloud platforms and Le Chat Pradeep Viswanathan / Neowin : Mistral announces Large 2 flagship LLM with 123 billion parameters Threads: @altstoreio : We've explicitly reviewed these initial sources to ensure they meet our safety standards.  Just add them from the Sources screen and new apps will appear in your Browse tab!  We've got a great first batch for y'all, keep reading to learn more 👇 X: Devendra Chaplot / @dchaplot : Super excited to announce Mistral Large 2 - 123B params - fits on a single H100 node - Natively Multilingual - Strong code & reasoning - SOTA function calling - Open-weights for non-commercial usage Blog: https://mistral.ai/... Weights: https://huggingface.co/... 1/N [image] Ravid Shwartz Ziv / @ziv_ravid : Mistral released a new 123B dense model, which is very good. Why is it dense and not a mixture of experts? There was a time a few months ago when it seemed that a mixture of experts was the promising direction, but it doesn't look like that anymore; why? LangChain / @langchainai : 🧠Mistral Large 2: improved reasoning and function calling 🧠 Try out the next generation of the @MistralAI flagship model, with its significantly improved reasoning, function calling, and more. Mistral announcement: https://mistral.ai/... LangChain PY docs: [image] Simon Willison / @simonw : Anyone know what benchmark @MistralAI are reporting on here for “Function Calling” for their new Mistral Large 2 model? Their post doesn't name the benchmark used: https://mistral.ai/... [image] @rajko_rad : What a week!!! it's amazing to have 405B open model but realistically you can't deploy that with in prod - here's the new SoTa deployable model 😎🔥 [image] Rowan Cheung / @rowancheung : Wow. After only one day of Llama 3.1 405b, French startup Mistral AI dropped LARGE 2. It's ANOTHER open-source flagship AI model that scores close to Llama 3.1 405b and even surpasses it on coding benchmarks while being much smaller at 123b. Benchmarks vs. Llama 3.1 405b: - [image] Richard Kelley / @richardkelley : Interesting to see “fits on a single GPU” as a selling point: @bilaltwovec : props to mistral for keeping w the efficiency corner branding after the last time [image] @teknium1 : btw now we have a 4th model in the public sphere comparable to gpt4. It's also like a spectrum of full open access to fully closed now lol Llama-3.1 - Fully Open (yea not fully open source) Mistral Large 2 - Open but non-commercial Sonnet 3.5 - closed GPT4 - closed [image] @teknium1 : Mistral just released a 123B param instruct tune only model called Mistral Large, unfortunately is a non-commercial license, and no base model was released, but looks great for personal use: Model: https://huggingface.co/... Blog Post: https://mistral.ai/... @guillaumelample : Today, we release Mistral Large 2, the new version of our largest model. Mistral Large 2 is a 123B-parameter model with a 128k context window. On many benchmarks (notably in code generation and math), it is superior or on par with Llama 3.1 405B. Like Mistral NeMo, it was trained @mistralai : https://mistral.ai/...

VentureBeat Shubham Sharma

Discussion

  • @altstoreio @altstoreio on threads
    We've explicitly reviewed these initial sources to ensure they meet our safety standards.  Just add them from the Sources screen and new apps will appear in your Browse tab!  We've got a great first batch for y'all, keep reading to learn more 👇
  • @dchaplot Devendra Chaplot on x
    Super excited to announce Mistral Large 2 - 123B params - fits on a single H100 node - Natively Multilingual - Strong code & reasoning - SOTA function calling - Open-weights for non-commercial usage Blog: https://mistral.ai/... Weights: https://huggingface.co/... 1/N [image]
  • @ziv_ravid Ravid Shwartz Ziv on x
    Mistral released a new 123B dense model, which is very good. Why is it dense and not a mixture of experts? There was a time a few months ago when it seemed that a mixture of experts was the promising direction, but it doesn't look like that anymore; why?
  • @langchainai LangChain on x
    🧠Mistral Large 2: improved reasoning and function calling 🧠 Try out the next generation of the @MistralAI flagship model, with its significantly improved reasoning, function calling, and more. Mistral announcement: https://mistral.ai/... LangChain PY docs: [image]
  • @simonw Simon Willison on x
    Anyone know what benchmark @MistralAI are reporting on here for “Function Calling” for their new Mistral Large 2 model? Their post doesn't name the benchmark used: https://mistral.ai/... [image]
  • @rajko_rad @rajko_rad on x
    What a week!!! it's amazing to have 405B open model but realistically you can't deploy that with in prod - here's the new SoTa deployable model 😎🔥 [image]
  • @rowancheung Rowan Cheung on x
    Wow. After only one day of Llama 3.1 405b, French startup Mistral AI dropped LARGE 2. It's ANOTHER open-source flagship AI model that scores close to Llama 3.1 405b and even surpasses it on coding benchmarks while being much smaller at 123b. Benchmarks vs. Llama 3.1 405b: - [imag…
  • @richardkelley Richard Kelley on x
    Interesting to see “fits on a single GPU” as a selling point:
  • @bilaltwovec @bilaltwovec on x
    props to mistral for keeping w the efficiency corner branding after the last time [image]
  • @teknium1 @teknium1 on x
    btw now we have a 4th model in the public sphere comparable to gpt4. It's also like a spectrum of full open access to fully closed now lol Llama-3.1 - Fully Open (yea not fully open source) Mistral Large 2 - Open but non-commercial Sonnet 3.5 - closed GPT4 - closed [image]
  • @teknium1 @teknium1 on x
    Mistral just released a 123B param instruct tune only model called Mistral Large, unfortunately is a non-commercial license, and no base model was released, but looks great for personal use: Model: https://huggingface.co/... Blog Post: https://mistral.ai/...
  • @guillaumelample @guillaumelample on x
    Today, we release Mistral Large 2, the new version of our largest model. Mistral Large 2 is a 123B-parameter model with a 128k context window. On many benchmarks (notably in code generation and math), it is superior or on par with Llama 3.1 405B. Like Mistral NeMo, it was trained
  • @mistralai @mistralai on x
    https://mistral.ai/...