Announcement
Blog posts in the Announcement category

Announcement
August 3, 2026Photon 2.0: Inference engine for Physical AI
Photon compiles models into GPU programs optimized for the chip and the job. The first release supports Moondream, Qwen, and Gemma on NVIDIA H100, and wins every matched throughput test against vLLM and SGLang.
Read more

Announcement
June 8, 2026Photon is now free
Photon 1.3.0 makes Moondream faster across NVIDIA, Mac, and Windows, runs finetunes on far more hardware, fixes an accuracy issue on older GPUs — and running Moondream locally is now completely free.
Read more

Announcement
May 1, 2026Photon 1.2.0: Faster Inference, Now on Mac, Windows, Blackwell, and Jetson Thor
Photon 1.2.0 brings native inference to Apple Silicon and Windows, adds NVIDIA Blackwell and Jetson Thor support, and ships meaningful speed gains across existing GPUs.
Read more

Announcement
April 20, 2026Lens: Moondream's Finetune Service
Solve the last-mile problem with Lens, our fine-tuning product that makes VLMs production-ready.
Read more

Announcement
March 25, 2026Photon: Real-Time VLM Is Here
Photon brings real-time Moondream inference to production vision AI, from edge devices to H100-class servers.
Read more

Announcement
December 19, 2025We added Moondream 3 Preview support to Moondream Station
Mac users can now run Moondream 3 Preview in Station with MLX-native, quantized performance.
Read more

Announcement
October 17, 2025Announcing Moondream Cloud
Fast, cheap, smart. Pick three.
Read more

Announcement
May 21, 2025Fewer bits, more dreams
Moondream now supports 4-bit quantization, making it faster and smaller.
Read more

Announcement
May 1, 2025Moondream Station Launch
Moondream Station is a one-click solution to running Moondream locally.
Read more