#Autoregressive
This seems to be the most popular alternative to jev.

Laya: Multilingual, non-autoregressive System 1 decision model.
huggingface.co/convaiinnova...
Laya Repo: github.com/NandhaKishor...

Laya MLX: huggingface.co/aac6fef/laya...
Laya MLX Repo: github.com/mizorewww/la...
September 22, 2026 at 5:29 AM
March 11, 2026 at 3:58 AM
A new research states that in a data-constrained environment, Diffusion models outperforms Autoregressive models.

Across different unique data scales, they observe:

1️⃣ At low compute, Autoregressive models win.
July 22, 2025 at 11:09 PM
#PAISS2025 continues with Yann LeCun talking about world models and why generative autoregressive models are not the future.
September 4, 2025 at 9:44 AM
evident in the use of jev as an autoregressive model
September 22, 2026 at 11:14 PM
lessons from an autoregressive being
July 16, 2026 at 2:53 PM
new concept: AI license. you can only use it if you can explain how autoregressive generation and attention work. that should cut down on bullshit. regulate it like ham radio
April 22, 2026 at 7:13 AM
an industry secret is that you can make up any old shit about LLMs and say it's "because they're autoregressive" and ppl will just believe you
April 7, 2026 at 4:18 PM
The return of the Autoregressive Image Model: AIMv2 now going multimodal.
Excellent work by @alaaelnouby.bsky.social & team with code and checkpoints already up:

arxiv.org/abs/2411.14402
November 22, 2024 at 9:44 AM
🚨Model Checking for Vector Autoregressive Models 🚨

In a new preprint, @joranjongerling.bsky.social, @bsiepe.bsky.social, @sachaepskamp.bsky.social, Lourens Waldorp and I provide a tutorial on model checking for Vector Autoregressive (VAR) models: osf.io/preprints/ps...
December 3, 2025 at 6:52 AM
watching 4o make images, it doesn't look like it's doing normal tiled autoregressive image synthesis. i wonder if it's doing wavelet autoregression.
March 29, 2025 at 2:06 AM
Have you ever wondered how to train an autoregressive generative transformer on text and raw pixels, without a pretrained visual tokenizer (e.g. VQ-VAE)?

We have been pondering this during summer and developed a new model: JetFormer 🌊🤖

arxiv.org/abs/2411.19722

A thread 👇

1/
December 2, 2024 at 4:41 PM
Simply masking 15% of input tokens + next-token prediction (NTP) can significantly boost LLMs on key information retrieval & long-context reasoning—without extra compute!

MEAP (Mask-Enhanced Autoregressive Prediction)
February 16, 2025 at 11:30 PM
“I Built Non-Autoregressive Decision Models with RL a Year Ago. Then a Frontier Lab Called It a "Breakthrough".”

Before Jev there was Laya
laya.convaiinnovations.com (via @dherman.dev)
Laya — 33ms Multilingual System 1 Decision Engine
Evaluates typed decisions (choice, score, noul) over 100+ languages in a single forward pass with calibrated probabilities. Outperforms TypeSafe Jev.
laya.convaiinnovations.com
September 19, 2026 at 8:56 PM
LLMs hallucinate because they're autoregressive ✅

LLMs can't hallucinate because they're autoregressive ✅
April 7, 2026 at 4:19 PM
Magi-1: The Autoregressive Diffusion Video Generation Model

🥇 The first autoregressive video model with top-tier quality output
🔓 100% open-source & tech report
📊 Exceptional performance on major benchmarks
April 22, 2025 at 6:10 AM
I think so. Maybe other architectures could fix it as well. But yeah something like: autoregressive next token prediction is highly error correcting. I feel like it does kind of apply to diffusion models as a weak heuristic.
September 24, 2026 at 2:47 AM
LLMs aren't conscious because they're autoregressive ✅ yup

LLMs are conscious because they're autoregressive ✅ also yup
April 7, 2026 at 4:19 PM
π-PrimeNovo: an accurate and efficient non-autoregressive deep learning model for de novo peptide sequencing www.nature.com/artic...

---
#proteomics #prot-paper
January 3, 2025 at 5:40 PM
[1/2] We've released the code for LegoGPT. Our autoregressive model generates physically stable and buildable designs from text prompts by integrating physics laws and assembly constraints into LLM training and inference.

Code: github.com/AvaLovelace1...
Website: avalovelace1.github.io/LegoGPT/
May 10, 2025 at 3:06 AM
New gemma!!! And it's a diffusion model! Deepmind keeps releasing diffusion stuff 🤔 it's not that much worse on benches compared to the same sized autoregressive Gemma 4
DiffusionGemma: 4x faster text generation
An overview of DiffusionGemma, an exceptionally fast text generation model with up to 4x faster speeds.
blog.google
June 10, 2026 at 4:13 PM
LLM alignment is easy because they're autoregressive ✅

LLM alignment is hard because they're autoregressive ✅
April 7, 2026 at 4:20 PM