Dubbed "Le Chonk" by its developers, the model is currently accessible only through a controlled public endpoint. Mistral plans to release the weights publicly in three weeks, following a rigorous safety evaluation process. Pierre Stock, the company's VP of Science, emphasized that this delay allows the team to collaborate with government partners, ensuring the technology supports defensive security applications rather than facilitating malicious activity.
The development of Mistral Large 4 marks a technical efficiency milestone. The model was trained using 4,000 NVIDIA GPUs, a fraction of the hardware typically utilized by major competitors in the United States and China. This lean approach to compute is expected to yield competitive performance in high-stakes sectors, specifically cybersecurity, finance, and semiconductor design. The focus on chip architecture aligns with the strategic interests of key backers, including Dutch lithography leader ASML and Samsung, the latter of which recently led a funding round that valued the French startup at approximately €21 billion.

Comments (0)
No comments yet. Be the first!