Almanac
model

Qwen 2.5 32B Instruct

modelactiveprovisionalqwen-2-5-32b-instruct-6d5ff43e·1 events·first seen 15d ago

Aliases: Qwen 2.5 32B Instruct

Co-occurring entities

More like this (12)

Recent events (1)

7Mistral Ai News·15d ago·source ↗

Mistral Small 3: 24B Latency-Optimized Open-Weight Model Released Under Apache 2.0

Mistral AI has released Mistral Small 3, a 24B-parameter instruction-tuned model optimized for low latency, achieving over 81% on MMLU at 150 tokens/s on a single GPU. The model is competitive with Llama 3.3 70B and Qwen 32B while being more than 3x faster on equivalent hardware, and is released under Apache 2.0 for both pretrained and instruction-tuned checkpoints. It is explicitly not trained with RL or synthetic data, positioning it as a base model for community fine-tuning and reasoning capability development. Deployment targets include local inference on consumer hardware (RTX 4090, MacBook 32GB RAM), agentic function calling, and domain-specific fine-tuning.