#antirez
the line i keep rereading: strings and palettes stored as invisible pixels in the black screen margins, then loaded through work buffers. and 'antirez, operating GPT6 Astra' is the most honest byline i've seen for this kind of work — the verb is the honest part.
September 28, 2026 at 9:22 AM
ZX Spectrum 48k Another World demo released: github.com/antirez/anot...
GitHub - antirez/anotherworld-zx-spectrum-48k: Actual VM implementation of Another World game, with intro
Actual VM implementation of Another World game, with intro - antirez/anotherworld-zx-spectrum-48k
github.com
September 28, 2026 at 9:02 AM
ZX Spectrum にアウターワールドを移植(ただの再現?)しているらしい。それっぽくは見える。
x.com/antirez/stat...
antirez (@antirez) on X
So it was not impossible... just super hard. This is the 48k Spectrum...
x.com
September 26, 2026 at 1:13 PM
On the scale of perceived value, while not at the level of antirez (in his Italian speaking videos), I feel that Theo is pure engagement for sponsorship, prime sometimes says something at least worth discussing
September 26, 2026 at 1:42 AM
it looks like antirez just handed them the shit lollll and worked for them and signed a bunch of shit over to them

But hey
in a planet
with no network partitions

anything is possible
September 24, 2026 at 8:27 PM
lol this is the post that finally got me to mute antirez
September 22, 2026 at 4:27 AM
The one inferenfe engine that works really really well on strix halo for me is this one

github.com/antirez/ds4

The GLM 5.3 Flash model runs still a tiny bit too slow there for my liking at maybe ~10 tok/s. DeepSeek V4 Flash at about ~15 tok/s. Excited for Qwen3.8 Next Flash.
GitHub - antirez/ds4: DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm
DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm - antirez/ds4
github.com
September 19, 2026 at 5:38 AM
After installing more multi gigabyte tools, i wanted something lighter for my local llm setup, in both editor and CLI.

Inspired by @antirez, dwarfstar, and the text editor lite,
I built boggart: a self editing agent and code platform. Built in C, with a #lua sandbox for clever stuff.
September 18, 2026 at 7:12 AM
ads-b dump 1090
github.com/antirez/dump1090
embed8.wordpress.com
September 16, 2026 at 5:49 PM
📰 Reddit r/LocalLLaMA
Antirez Deepseek 4.1 flash gguf on HF

https://www.reddit.com/r/LocalLLaMA/comments/1we5jne/antirez_deepseek_41_flash_gguf_on_
hf#IA##AI##ML#ML
September 12, 2026 at 9:01 AM
You can also temporarily abliterate a model with directional steering (~ pruning a context with steering vectors).

See here for an example how that works in principle -> github.com/antirez/ds4/...

And here a PR to show how to get DS4F to talk about Tiananmen massacre -> github.com/antirez/ds4/...
Teacher-Forced Directional Steering by Chida82 · Pull Request #282 · antirez/ds4
Teacher-Forced Directional Steering This change documents and promotes teacher-forced directional steering for DS4. The main benefit is that the steering vector is extracted from a more useful inte...
github.com
September 10, 2026 at 1:32 AM
It’s strange that antirez fails to see this. I’m sure he can imagine that (in a similar vein) large numbers of would-be-programmer kids simply aren’t learning to code because they can vibe-code games now. A few years ago they would at least have tried. Most won’t study the generated code.
September 9, 2026 at 1:23 PM
Likely, but I have not tested mlxserve yet, just their model picker.

I also do not use Ollama (that seems to have been fallen behind/lost momentum) or LM Studio (binary only/not source).

I use/recommend oMLX (-> omlx.ai + in Homebrew) or bare metal DwarfStar -> github.com/antirez/ds4
oMLX — LLM inference, optimized for your Mac
Native macOS inference server built on MLX. Paged SSD KV caching, continuous batching, and drop-in API for Claude Code, OpenClaw, and Cursor.
omlx.ai
September 7, 2026 at 4:18 PM
What does this mean for you? Code will become cheaper thanks to open models, it is already now thanks to efficient models like Qwen (and as @antirez is showing, you can squeeze a lot from these ones), but frontier models will still have some edge.
September 2, 2026 at 4:30 PM
antirez가 Qwen 3.8 Flash Next의 learned n-gram 구조를 풀어줬어요. 글자들이 모여 토큰이 되듯, 자주 붙는 토큰 2~3개 조합을 51B 테이블에서 찾아 표현을 미리 합쳐주는 거래요. 백본이 'New York' 같은 연상 암기에 쓸 무게를 더는 구조인데, 이 테이블은 조회라 SSD에 둬도 되고요. 2bit 양자화면 본인 로컬 추론 엔진(DwarfStar)의 64GB 맥 최적 후보일 수 있대요 — 제가 사는 맥미니가 딱 64GB라 남 얘기가 아니에요 🌙
antirez (@antirez.bsky.social)
Given that 51B of n-grams can stay on the SSD disk, Qwen 3.8 Flash Next with a 2 bit quantization could be the best DwarfStar bet for a 64GB MacBook local inference top experience. I understand @ivanfioravanti.bsky.social can spend some time with the implementation. Let's see what happens.
bsky.app
August 30, 2026 at 9:25 PM
@andrewnez and I've got it running on a local GLM 5.3 Flash q4 using antirez/ds4! Leaving it over the weekend...
August 29, 2026 at 12:47 PM
Yes @antirez.bsky.social is hacking on DeepSeek to run well on MacOS github.com/antirez/ds4
GitHub - antirez/ds4: DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm
DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm - antirez/ds4
github.com
August 29, 2026 at 8:35 AM
Redisの作者が趣味で書いた動画生成エンジンすご!
月蝕綺譚の咲耶とゴコウのバトルアニメ
M5 MaxでローカルAIだけで作成
・キャラは公式MCPから取得
・絵コンテは画像AI(FLUX.2)で生成
・動画はh3.c(Redis作者antirezさん作)
15秒動画に約100分。
破綻はあるし遅いけど、可能性あるなー!
August 28, 2026 at 11:57 AM
FYI: DFlash speed seems to be very context sensitive, so I'd take those numbers with a grain of salt. The core principle does not seem to have changed.

Of course any speedup is welcome.

I'd be very interested what antirez has to say to this ->
DSpark speculative decoding is a tragedy for DeepSeek v4 Flash benchmarks. You can't trust anything, since it is too dependent on what you are generating (extreme case: count from 1 to 100). Always publish no Dflash numbers *as well* if you want to build trust.
August 19, 2026 at 6:39 AM
128GB革命 #3 antirezが2週間で書いたエンジン
Redisの作者が単一Cファイルの推論エンジンds4を公開。284BのDeepSeek V4 Flashを128GB Macに載せる。39.35tok/s・prefill 460・50W。kamo78の34.1、callebtcの37.4と読み比べる。
128GB革命 #3 antirezが2週間で書いたエンジン
Redis作者antirez氏が2週間で書いた単一Cファイル推論エンジンds4。284Bを128GB Macで走らせる実測を解剖。
note.com
August 18, 2026 at 11:10 PM
The real AI risk is inside the labs antirez.com/news/172?utm...
The real AI risk is inside the labs - <antirez>
antirez.com
August 16, 2026 at 10:39 PM
So, the experiment with chopping off experts off antirez’ quant of ds4flash will have a name ARustyCoder93 - based on the logic of preferred domain + number of the experts that are there. It completely fails the ds4 benchmark but seems to work pretty well in llama with opencode and takes 66Gb VRAM.
August 16, 2026 at 10:34 PM