Mistral Large 4: "Le Chonk"

(mistral.ai)

354 points | by j-bu 2 hours ago

25 comments

  • xpct 1 hour ago
    I don't know why, but I personally find Mistral's marketing strategy much more appealing than that of other companies.

    For example, there's something about Anthropic's picked design and their little Claude avatars that's unsettling to me.

    • booty 25 minutes ago
      Sans context, I really like the Anthropic and Claude's faux-academic, minimal-ish, intellectual-ish branding.

      (Especially the way it looked ~12 months ago -- it's gotten more cluttered since then. Perhaps unavoidably, as the breadth of their offerings has grown)

      But over time it's begun to feel like unsettling cognitive dissonance as their ambitions grow and the stuff to worry about has piled up.

    • isoprophlex 27 minutes ago
      The logo isnt a stylized butthole. That sure helps endear me to them.
    • fidotron 1 hour ago
      The entire Anthropic branding is religious kitsch - deeply off putting, but apparently quite reflective of their reality.
    • oytis 40 minutes ago
      And their cookie banner. Never thought I would like a cookie banner
    • walrus01 1 hour ago
      I think there's such a thing as throwing too many marketing and sales people and too much polishing and "refinement" at something. I see in new product announcements from Microsoft as well. It's like seeing someone try too hard to impress you.
  • vessenes 1 hour ago
    Very nice progress. Also I respect putting Kimi on those charts. Regardless of if they are beating the Pareto frontier (not now), model diversity is a good thing for humanity — I’m hopeful for the team to keep increasing their gains.
  • chevman 1 hour ago
    This is awesome, one of the coolest Pokemon ever too for those that don't follow that universe :)

    https://bulbapedia.bulbagarden.net/wiki/Lechonk_(Pok%C3%A9mo...

    • maelito 15 minutes ago
      Was wondering what Le Chonk meant (not French).
  • danslo 1 hour ago
    Open weight, European, competes with GLM-5.3 on cybersecurity. What's not to like?
  • DevKoala 26 minutes ago
    One step closer to Le Chaton Fat.
  • juliennakache 11 minutes ago
    Is that a reference to LeChuck in Monkey Island? Love that game!
  • rglover 1 hour ago
    Excited to try this. The low costs v. benchmarks alone here are worth a serious test. K3 has been my daily driver for a month or two now and it's dramatically reduced token spend (while not having much of a negative impact on productivity).

    This was the era of the AI race I was waiting for.

    • netvarun 43 minutes ago
      Curious now that SOL is cheaper than k3 - is k3 still your primary workhorse?
  • bartstp 52 minutes ago
    How could I resist switching to a model named after my cat!?
    • laserbeam 48 minutes ago
      Stats be damned irrelevant. The naming is good with this one!
    • volkk 15 minutes ago
      you and 10k other redditors
  • amelius 1 hour ago
  • timcobb 38 minutes ago
    > Trained from scratch

    How are they training without pirating the Z library corpus and all that?

    • sigmar 5 minutes ago
      I interpret "scratch" to mean brand new weights. Not that they aren't training on a corpus of human text
  • taspeotis 1 hour ago
    • stronglikedan 1 hour ago
      Of course not. That's clearly a different URL.
  • aennassiri 1 hour ago
    Excited to see this! Nice that they are saying this is just a first step.

    Give them more compute!

  • maz1b 1 hour ago
    I'm glad they're keeping at it!
  • spwa4 2 hours ago
    Don't believe Mistral. They're wrong. It's really called "Le chaton fat".

    Also 1T-A49B. Weights currently closed but promise to open source them by the end of the month.

    Great release movie.

    • alterom 1 hour ago
      >Don't believe Mistral. They're wrong. It's really called "Le chaton fat".

      OpenAI's therapist: Le Chaton Fat isn't real and cannot hurt you

      Le Chaton Fat:

  • tosh 1 hour ago
    sorting the charts like that gives off weird vibes
    • jasonjmcghee 1 hour ago
      Had the same thought - feels chart crime adjacent
    • lern_too_spel 1 hour ago
      Bar charts should start at zero. If they don't start at zero, there should be a clear visual indicator that the chart has been trimmed without having to read the axis labels. I hate that this has to be repeated so often that it has become a cliché.
  • redanddead 1 hour ago
    better than K3 and DS4, cool
  • ChrisArchitect 1 hour ago
  • Razengan 59 minutes ago
    Awh I was half expecting a zombie pirate..
  • theturtletalks 1 hour ago
    A bit disappointing to see it still lagging behind Chinese open models. Those Chinese models are pushing proprietary models to raise the bar, but we need equally strong non-Chinese open models to challenge the Chinese ones in turn.
  • plastic-enjoyer 1 hour ago
    We are so back
  • clavicle1009 1 hour ago
    YES finally
  • baggachipz 1 hour ago
    Now THAT'S how you name a model. Take note, others.
  • __natty__ 1 hour ago
    [flagged]
  • catlover76 51 minutes ago
    [dead]
  • sinan-faizal 59 minutes ago
    i subbmitted a partnership proposal in your contact.