-rough sketch-
Illustration for Moondream novella
-rough sketch-
Illustration for Moondream novella
www.youtube.com/watch?v=T7sx...
www.youtube.com/watch?v=T7sx...
(Rough sketch for my as-yet unpublished illustrated novella Moondream)
(Rough sketch for my as-yet unpublished illustrated novella Moondream)
they have a demo, but how do you execute vision models locally?
they have a demo, but how do you execute vision models locally?
Try it out here: moondream.ai/playground
Try it out here: moondream.ai/playground
Live demo: huggingface.co/spaces/moond...
Blog post: moondream.ai/blog/announc...
Live demo: huggingface.co/spaces/moond...
Blog post: moondream.ai/blog/announc...
Runs Moondream, Qwen 3.5, and Gemma 4 on NVIDIA H100. Photon outperformed vLLM and SGLang in every throughput test and brought every model online faster, via use of a single 'megakernel' that runs the entire inference on the GPU alone.
Runs Moondream, Qwen 3.5, and Gemma 4 on NVIDIA H100. Photon outperformed vLLM and SGLang in every throughput test and brought every model online faster, via use of a single 'megakernel' that runs the entire inference on the GPU alone.
A 9B param, 2B active MoE vision hybrid reasoning vision language model that supports both reasoning and non-reasoning mode. It focus on visually grounded reasoning, where the model references objects and spatial positions in the image while doing said reasoning,
A 9B param, 2B active MoE vision hybrid reasoning vision language model that supports both reasoning and non-reasoning mode. It focus on visually grounded reasoning, where the model references objects and spatial positions in the image while doing said reasoning,
You can run Moondream vision-language models, object detection, image segmentation (SAM 3), and even train your own geospatial segmentation model end-to-end.
You can run Moondream vision-language models, object detection, image segmentation (SAM 3), and even train your own geospatial segmentation model end-to-end.