Skip to main content
ToolPotion

SmolLM3 Language Model

SmolLM3 is a 3B parameter language model designed to push the boundaries of small models. It supports dual mode reasoning, six languages, and long context processing, offering strong performance at the 3B–4B scale.

View Model
Share

Description

SmolLM3 is a 3 billion parameter language model developed to advance the capabilities of small models in artificial intelligence. It is designed to support dual mode reasoning and is multilingual, natively supporting six languages: English, French, Spanish, German, Italian, and Portuguese. The model is fully open, with open weights and comprehensive training details, including public data mixtures and training configurations.

The architecture of SmolLM3 is a decoder-only transformer using GQA and NoPE with a 3:1 ratio. It was pretrained on 11.2 trillion tokens using a staged curriculum that included web, code, math, and reasoning data. Post-training involved midtraining on 140 billion reasoning tokens, followed by supervised fine-tuning and alignment through Anchored Preference Optimization (APO).

SmolLM3 is optimized for hybrid reasoning and supports long context processing, trained on a 64k context and capable of handling up to 128k tokens using YARN extrapolation. The model can be deployed using vLLM and SGLang, compatible with OpenAI format APIs. It also supports tool calling, allowing integration with various tools through XML or Python function calls.

For local inference, SmolLM3 can be used with llama.cpp, ONNX, MLX, MLC, and ExecuTorch. Quantized checkpoints are available for efficient deployment. The model's performance has been evaluated across various benchmarks, demonstrating strong results in multilingual Q&A, competitive programming, and instruction following tasks.

SmolLM3 is a valuable tool for developers and researchers looking to leverage AI for multilingual and complex reasoning tasks. However, users should be aware that the generated content may not always be factually accurate or free from biases present in the training data.

SmolLM3 Language Model Highlights

  • 3B parameter language model

  • Supports dual mode reasoning

  • Multilingual: 6 languages

  • Long context processing up to 128k tokens

  • Open model with open weights

  • Instruct model optimized for hybrid reasoning

  • Tool calling support

  • Compatible with vLLM and SGLang

Getting Started with SmolLM3 Language Model

  1. Access page: Visit the Hugging Face model page

  2. Load model: Use transformers v4.53.0 or latest vllm

  3. Configure environment: Set sampling parameters and context length

  4. Integrate: Use tool calling with XML or Python functions

  5. Fine-tune: Apply supervised fine-tuning and alignment

SmolLM3 Language Model's Use Cases

  • Multilingual Q&A
  • Hybrid Reasoning
  • Tool Integration
  • Long Context Processing
  • Instruction Following

FAQ from SmolLM3 Language Model

From Hugging Face

SmolLM3 Language Model Reviews

Loading...

Popular AI Tools Like SmolLM3 Language Model

Qwen3-8B is a large language model designed for advanced reasoning, instruction-following, and multilingual support. It features seamless mode switching for optimal performance in…

FeaturedAI Models & LLMs

Gemma 3-27B IT is a state-of-the-art multimodal AI model from Google, capable of handling text and image inputs to generate text outputs. It supports over 140 languages and is…

AI Models & LLMs

Falcon3-10B-Instruct is a state-of-the-art language model designed for reasoning, language understanding, and instruction following tasks. It supports multiple languages and…

AI Models & LLMs

AI Models

Hy3 Preview by Tencent Hunyuan is an open-source language model featuring 295 billion parameters. It excels in math, coding, and multilingual tasks, offering competitive…

AI Models & LLMs

Kimi K3 is an advanced open-weight multimodal AI model designed for long-horizon coding, knowledge work, and reasoning. It features a 1-million-token context window and is built…

FeaturedAI Models & LLMs

Phi-4-reasoning-plus is a state-of-the-art reasoning model developed by Microsoft Research. It is fine-tuned from Phi-4 using supervised learning and reinforcement learning,…

AI Models & LLMs

Mistral Small 4 is a versatile AI model that unifies reasoning, coding, and multimodal capabilities into a single platform. It allows users to customize, fine-tune, and deploy AI…

FeaturedAI Models & LLMs

Llama-3.1-8B-Instruct is a multilingual large language model developed by Meta, optimized for instruction-based tasks. It is designed for commercial and research applications,…

FeaturedAI Models & LLMs