StartupXO
Search
English

Find XO: Mistral Large 4 is an AI model with 490 billion parameters that handles both text and images (docs.mistral.ai)

No votes yet startupxo 1 comment

Writing language: Korean Read in the original language

Summary / Read source ↗

- Mistral Large 4 is a multimodal AI model with 490 billion active parameters and a total of 1 trillion 500 billion parameters (source: Mistral AI official documentation). - It includes a visual encoder consisting of 1 hundred million 6 million parameters, enabling it to process not only text but also image inputs. - The API supports various interactive features such as chat, document Q&A, function calling, and batch processing. - Usage costs are charged based on the number of tokens, with processing input 100 tokens costing about 0.68 dollars (according to the official pricing).
Found on

Hacker News ↗ / 828 votes / 484 comments

Sign in to comment

1 comment

Editorial opinion startupxo

It's notable that the visual encoder is separately integrated at a scale of 1.6B parameters, which sets it apart from simply extending to text. I'm curious about how processing delays or cost increases occur during multimodal input handling, and if the official documentation includes data on response times under real-world usage conditions or the actual cost impact when handling large volumes, that would be very helpful for making adoption decisions.
Writing language: Korean

Keyboard shortcuts

Choose a post with the up and down arrows, then press Enter.

↑ / ↓
Previous post / next post
Enter
Open summary and comments for the selected post
Tab
Move to the submit or comment button, then press Enter

Type normally in text fields. Tab and Enter are always available.