Ternary Bonsai 2 (Qwen3.8-27B) at ~1200-1750tps prefill, ~100tps (peak ~400tps) decode on a 4080 super
Ternary Bonsai 2 (Qwen3.8-27B) at ~1200-1750tps prefill, ~100tps (peak ~400tps) decode on a 4080 super
$12k Mac Studio (2x 256GB or 1x 512GB)
$14k DGX GB-10 Spark (4x)
$14k AMD Strix cluster (4x)
$35k AMD Mi210 (8x)
$55k NVIDIA RTX6000 PRO (6x)
$12k Mac Studio (2x 256GB or 1x 512GB)
$14k DGX GB-10 Spark (4x)
$14k AMD Strix cluster (4x)
$35k AMD Mi210 (8x)
$55k NVIDIA RTX6000 PRO (6x)
最低RTX4000、推奨がRTX6000当たりに合わせてくるのかな。
最低RTX4000、推奨がRTX6000当たりに合わせてくるのかな。
context length 32768
nate@RTX6000:~$ ollama show gemma3:270m | grep context
context length 32768
nate@RTX6000:~$ ollama show gemma3:27b | grep context
context length 131072
nate@RTX6000:~$
context length 32768
nate@RTX6000:~$ ollama show gemma3:270m | grep context
context length 32768
nate@RTX6000:~$ ollama show gemma3:27b | grep context
context length 131072
nate@RTX6000:~$
Built for teams working on AI inference, model testing, rendering, simulation and visual production on European cloud infrastructure.
www.exoscale.com/gpu/rtx6000/
#gpu #ai
Built for teams working on AI inference, model testing, rendering, simulation and visual production on European cloud infrastructure.
www.exoscale.com/gpu/rtx6000/
#gpu #ai
👉 1 * RTX6000 Ada 48 GB for $1.50/hour!!!
Book here: gpucompare.com/clusters/opp...
#cloud #ai #machinelearning #gpu
👉 1 * RTX6000 Ada 48 GB for $1.50/hour!!!
Book here: gpucompare.com/clusters/opp...
#cloud #ai #machinelearning #gpu
Other thought:
How many chickens can each $10,000 RTX6000 Pro buy to feed us?
Other thought:
How many chickens can each $10,000 RTX6000 Pro buy to feed us?
Quadroだった
Quadroだった
"Sapphire only makes AMD GPU's. They mostly use Nvidia RTX6000 ADA and RTX5090's."
"Sapphire only makes AMD GPU's. They mostly use Nvidia RTX6000 ADA and RTX5090's."
tok/watt is better on Macs.. and software is easier these days.. MLX will improve
and IMO: nvidia is ripe for competitors.. and I think Apple is better positioned than AMD for consumers
tok/watt is better on Macs.. and software is easier these days.. MLX will improve
and IMO: nvidia is ripe for competitors.. and I think Apple is better positioned than AMD for consumers
I don't develop on it directly. Everything goes through ssh/tailscale to my big home machine, 9995wx/768gb/96gb rtx6000 pro Blackwell. I can't lug that compute with me.
Spending a lot of money specifically on a fast *laptop* seems like a bad idea to me.
I don't develop on it directly. Everything goes through ssh/tailscale to my big home machine, 9995wx/768gb/96gb rtx6000 pro Blackwell. I can't lug that compute with me.
Spending a lot of money specifically on a fast *laptop* seems like a bad idea to me.
👉 1 * RTX6000 Ada 48 GB for $1.50/hour!!
Book here:https://gpucompare.com/providers/linode
#cloud #ai #machinelearning #gpu
👉 1 * RTX6000 Ada 48 GB for $1.50/hour!!
Book here:https://gpucompare.com/providers/linode
#cloud #ai #machinelearning #gpu
#NVIDIA #RTX6000 #PCGaming #IA
www.justjeuxvideo.com/nvidia-boule...
#NVIDIA #RTX6000 #PCGaming #IA
www.justjeuxvideo.com/nvidia-boule...
Training via mirror self-play on 3× RTX6000 Ada (~12h) led to strong gains vs. google/gemini-2.0-flash-lite-001—both in-domain (SimpleTak) and OOD (KuhnPoker). Learned strategies generalize, despite never playing vs. Gemini during training.
Training via mirror self-play on 3× RTX6000 Ada (~12h) led to strong gains vs. google/gemini-2.0-flash-lite-001—both in-domain (SimpleTak) and OOD (KuhnPoker). Learned strategies generalize, despite never playing vs. Gemini during training.