
Pixtral 12B 24.09Atvērts multimodāls 12B modelis, kas apstrādā iekavētus attēlus un tekstu ar 128K konteksta logu.
Pārskats
Galvenās funkcijas
- 12B parametru redzes‑valodas modelis
- Iekavēti attēlu un teksta ievades
- 128K tokenu konteksta garums
- Nativs mainīgo attēlu izmēru atbalsts
- Atvērta svara izlaidums
- Piemērots OCR, VQA un uzrakstu ģenerēšanai
Cenas
- Modelis
- Free
- Kategorija
- LLM
- Vērtējums
- 4.6 / 5 (5)
Lietošanas gadījumi
Multimodāla domāšana
Pixtral 12B spēj saprast gan dabiskas attēlus, gan dokumentus, sasniedzot šobrīd augstākos rezultātus MMMU domāšanas testā un pārsniedzot lielākos modeļus.
Instruāciju izpilde
Pixtral 12B izceļas instrukciju izpildē, it īpaši multimodālās un tikai teksta situācijās, ar 20 % salīdzinošu uzlabojumu teksta IF‑Eval un MT‑Bench salīdzinājumā ar tuvāko atvērtās koda modeli.
Multimodāla jautājumu atbildēšana
Pixtral 12B parāda spēcīgas spējas multimodālajā jautājumu atbildēšanā, tostarp dokumentu jautājumu atbildēšanā un diagrammu un attēlu saprašāšanā.
Plusi un mīnusi
Plusi
- Atvērtas svaras iespējas pašizvietošanai
- Apstrādā vairākus attēlus vienā pieprasījumā
- Liels 128K konteksta logs
- Fleksīvi attēlu izšķirtspējas un aspektu attiecības
Mīnusi
- Prasa ievērojamas GPU resursus
- Mazāk nekā modernākie slēptie modeļi
- Ierobežota rīku izvēle salīdzinājumā ar īpašnieku API
Kauju rekords
1 kaujā Panteonā.
Last battle
Atsauksmes
Vidējais no 5 vērtējumiem.
Pieslēdzies, lai atstātu atsauksmi.
Does the job
Pretty happy overall. Open-weight release just works and large 128K context window. but no dealbreakers — I'd recommend it to a friend without hesitating.
Does the job
Pretty happy overall. Open-weight release just works and handles multiple images per prompt. but no dealbreakers — I'd recommend it to a friend without hesitating.
Years in this space
I've evaluated a lot of these over the years. What stands out here is interleaved image and text inputs — handled better than most — and handles multiple images per prompt. Smaller than frontier closed models is my one real gripe. Worth the time if this is your use case.
Solid for our team
We rolled this out across the team last quarter and open weights for self-hosting. Open-weight release fits neatly into how we already work, and interleaved image and text inputs removed a step we used to do by hand. Smaller than frontier closed models, which is the main caveat, but it has held up under daily use.
Years in this space
I've evaluated a lot of these over the years. What stands out here is 12B parameter vision-language model — handled better than most — and open weights for self-hosting. Smaller than frontier closed models is my one real gripe. Worth the time if this is your use case.
Jautājumi
How many images and how much text can I include in a single prompt?
Pixtral supports interleaved image and text inputs within a single 128 K token context, allowing any number of images (at their natural resolution) alongside long‑form text in one prompt.
Asked by Lorenzo Bianchi · Nov 20, 2025
Is Pixtral 12B still maintained, and are there newer alternatives?
Pixtral 12B is deprecated and no longer maintained; Mistral AI recommends using its newer, more powerful vision‑language models that supersede Pixtral for production use.
Asked by Carlos Mendoza · Nov 10, 2025
What hardware is needed to run Pixtral 12B effectively?
The model requires substantial GPU memory due to its 12 billion parameters and 400 M‑parameter vision encoder; typical deployments use high‑end GPUs (e.g., A100 40 GB or comparable) to handle the 128 K token context and multiple images.
Asked by Ivo Novotný · Nov 6, 2025
Can I self‑host Pixtral 12B, and under what license?
Yes, Pixtral 12B is released under the Apache 2.0 open‑source license, allowing you to download the weights and run the model locally on your own hardware.
Asked by Priya Nair · Oct 12, 2025
Uzdod jautājumu
LLM alternatīvas

Vidvēlīgs LLM ģitve unificējot vairāk kā 1000 modeļus aiz vienas API

Nākotnes paaudzes, koncentrēts uz loģisko domāšanu AI modelis no DeepSeek

Atvērtā pirmkoda "mixture-of-experts" modelis, kas piedāvā GPT-4o līmeņa izpēti par daļi no izdevumiem.

Sarunārais AI no xAI, izstrādāts racionēšanai, pētniecībai un reāla laika atbildēm.

Meta daudzvalodu atvērtās svara LLM, pielāgota efektīvai, augstas kvalitātes teksta ģenerācijai.

AI balstīts MP3 uz tekstu pārveidotājs, kas pārvērš audio par tīru, lasāmu transkripciju.

Atvērts lielais valodas modelis, kas izceļas loģikas, matemātikas un kodēšanas uzdevumos, ar MIT licence brīvas izmantošanas un modificēšanas tiesībām.

OpenAI modelis, kas koncentrējas uz argumentāciju un ir izstrādāts sarežģītu, daudzpakāpju problēmu risināšanai.
Trending now

Precīza Pildījuma Palīdzība ar Vispārējām Izskaidrojumiem

Dokumentu intelekta API, kas parse, dalās, OCR un izvelk strukturētus datus no kompleksām PDF, slaidēm un kalkulāciju tabulām.

Finansētas atbildes, maksām uz klikšķi.

Līderu platforma automācijai procesu darbības, izmantojot autonomas mācības kādas, lai nodrošinātu darbības plūsmu šādaīs dažāda veida rūpīgajās nozarēs.
