So in the meantime have some juicy details of a strawberry. #3dgs
So in the meantime have some juicy details of a strawberry. #3dgs
📄 www.biorxiv.org/content/10.1...
💾 mmseqs.com
🗞️ developer.nvidia.com/blog/boost-a...
📄 www.biorxiv.org/content/10.1...
💾 mmseqs.com
🗞️ developer.nvidia.com/blog/boost-a...
Try it out here: moondream.ai/playground
Try it out here: moondream.ai/playground
replicate.com/fofr/any-com...
Here's an example:
replicate.com/p/6a8ydnzcqs...
15s on an L40S, out of the box.
replicate.com/fofr/any-com...
Here's an example:
replicate.com/p/6a8ydnzcqs...
15s on an L40S, out of the box.
They offload the key-value cache to host memory during inference, significantly reducing GPU memory pressure. As a result, it processes up to 3m tokens on a L40s 48GB - 3x larger.
Paper: arxiv.org/abs/2502.08910
They offload the key-value cache to host memory during inference, significantly reducing GPU memory pressure. As a result, it processes up to 3m tokens on a L40s 48GB - 3x larger.
Paper: arxiv.org/abs/2502.08910
www.mediamarkt.nl/nl/product/_...
www.mediamarkt.nl/nl/product/_...
Nvidia L40S GPUs now up on Replicate:
replicate.com/blog/nvidia-...
We've measured that on average they're about 40% faster than our A40s (depending on the model). Switch your models from A40 to L40S to get faster outputs, cheaper.
Nvidia L40S GPUs now up on Replicate:
replicate.com/blog/nvidia-...
We've measured that on average they're about 40% faster than our A40s (depending on the model). Switch your models from A40 to L40S to get faster outputs, cheaper.
- 91.6 TFLOPS (FP32)
- 733 TFLOPS (FP16)
- 1,466 TFLOPS (FP8)
- AMD EPYC 7313
- 256GB RAM
- Up to 4x GPUs (e.g. Nvidia L40S)
- 246TB NVMe storage
- 100GbE + 25GbE networking
- 2.5kW PSU, rugged case
- <55 lbs, fits in carry-on
- 91.6 TFLOPS (FP32)
- 733 TFLOPS (FP16)
- 1,466 TFLOPS (FP8)
- AMD EPYC 7313
- 256GB RAM
- Up to 4x GPUs (e.g. Nvidia L40S)
- 246TB NVMe storage
- 100GbE + 25GbE networking
- 2.5kW PSU, rugged case
- <55 lbs, fits in carry-on
replicate.com/fofr/any-com...
- flux canny dev (+lora)
- flux depth dev (+lora)
- flux redux dev
Here's a working example:
replicate.com/p/9q93d0f1bh...
This model also now runs on a faster L40S GPU.
replicate.com/fofr/any-com...
- flux canny dev (+lora)
- flux depth dev (+lora)
- flux redux dev
Here's a working example:
replicate.com/p/9q93d0f1bh...
This model also now runs on a faster L40S GPU.
It takes 60s to make 5 images, instead of 3 minutes.
Costing ~$0.06 instead of ~$0.13
Faster and cheaper 😅
replicate.com/fofr/consist...
It takes 60s to make 5 images, instead of 3 minutes.
Costing ~$0.06 instead of ~$0.13
Faster and cheaper 😅
replicate.com/fofr/consist...
Im just using ollama run X —verbose for testing, nothing too rigorous
Im just using ollama run X —verbose for testing, nothing too rigorous
[Prime] $449.99 | dreame L40s Ultra Robot Vacuum and Mop: $449.99!
Snag it now: #ad #deal
[Prime] $449.99 | dreame L40s Ultra Robot Vacuum and Mop: $449.99!
Snag it now: #ad #deal
amazon.com/dp/B0DYSWW2LN
amazon.com/dp/B0DYSWW2LN
479 €
#BonPlan #HighTech #Promo
l40s: is that plural?
rm -rf ~/
l40s: is that plural?