Home icon

langcache-embed-v3-small, Mellum2-12B-A2.5B-Thinking, and LightOnOCR-2-1B models now available on Amazon SageMaker JumpStart

News



This article announces the availability of three new foundation models on Amazon SageMaker JumpStart: langcache-embed-v3-small for semantic caching, Mellum2-12B-A2.5B-Thinking for code reasoning, and LightOnOCR-2-1B for document OCR.

  • langcache-embed-v3-small optimizes semantic caching by mapping queries to dense vectors, reducing redundant LLM calls and accelerating response times
  • Mellum2-12B-A2.5B-Thinking uses Mixture-of-Experts architecture for code generation, debugging, and multi-step reasoning with 131,072-token context length
  • LightOnOCR-2-1B provides multilingual document-to-text conversion for PDFs and images without brittle OCR pipelines
  • Deploy models with a few clicks via SageMaker JumpStart console or Python SDK

These models expand SageMaker JumpStart's foundation model portfolio, enabling customers to deploy specialized AI solutions for caching, coding, and document processing on AWS infrastructure.



Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

Aug 11
2026
FLUX.2-small-decoder and gemma-4-12B-it models now available on Amazon SageMaker JumpStart
Aug 11
2026
LocateAnything-3B, Qwen-AgentWorld-35B-A3B, and Qwen3.5-122B-A10B models now available on Amazon SageMaker JumpStart
Aug 11
2026
GLM-5.2 FP8, NVIDIA-Nemotron-Nano-12B-v2 and GLM-OCR models now available on Amazon SageMaker JumpStart
Aug 11
2026
NVIDIA Nemotron 3.5 Lightning model is now available on Amazon SageMaker JumpStart

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.