#TimK
Why did the Evil Ass Company™ copy Astro Thornback again?
#megamanDO #casualtiesunknown
#art #deltarot
September 25, 2026 at 3:12 PM
So...um...this happened. bsky.app/profile/timk...
September 22, 2026 at 11:55 PM
Superintelligence != Super Intelligence
September 22, 2026 at 5:14 PM
it writes fine
September 22, 2026 at 4:44 PM
Mathstra (OpenAI) has solved 100 open problems across most branches of math in the last 3 weeks

openai.com/index/adviso...
September 21, 2026 at 9:19 PM
Mathstra (OpenAI) has solved 100 open problems across most branches of math in the last 3 weeks

openai.com/index/adviso...
September 21, 2026 at 9:18 PM
they’re getting nationalized in a different way

bsky.app/profile/timk...
Trump *really* doesn’t like Anthropic
September 21, 2026 at 6:25 PM
UPDATE: huge props to @typesafeai.bsky.social

On Sept 19 they updated their terms to remove the sketchy clauses

- no longer require permission to build products on them
- clarified MCA vs ToS conflicts
- no benchmarking restriction

fast reaction time
don’t use Jev at work btw. it’s strictly and only for personal use and you have to ask permission to use it for professional use cases

ChatGPT | original

typesafe.ai/legal/terms
September 21, 2026 at 1:32 PM
UPDATE: huge props to @typesafeai.bsky.social

On Sept 19 they updated their terms to remove the sketchy clauses

- no longer require permission to build products on them
- clarified MCA vs ToS conflicts
- no benchmarking restriction

fast reaction time
don’t use Jev at work btw. it’s strictly and only for personal use and you have to ask permission to use it for professional use cases

ChatGPT | original

typesafe.ai/legal/terms
September 21, 2026 at 1:29 PM
a lot of people are thinking agents aren’t much different

an agent is a long running application like a web server, with potentially demanding memory & CPU requirements

except it only makes sense to have one instance of the “server”. Per user. And not waste resources. And app memory is durable
September 20, 2026 at 8:34 PM
no they're public too
bsky.app/profile/timk...
ok, maybe it’s less bad; the MCA walks back those clauses and gives permission to use for any purpose.

The MCA seems to mostly take precedence over the ToS. Although there’s some rough edges

typesafe.ai/legal/mca
Master customer agreement - TypeSafe AI
TypeSafe AI is an AI lab building machine-native intelligence infrastructure for automation, designed to make decisions within software. Try our first System One Model, Jev, in early access.
typesafe.ai
September 20, 2026 at 1:29 AM
yes, Jev got very hyped. Yes, overhyped

Jev has no easy moat, all labs will have an equivalent in a few months. Their only moat is brand recognition

honestly can’t blame them
“I Built Non-Autoregressive Decision Models with RL a Year Ago. Then a Frontier Lab Called It a "Breakthrough".”

Before Jev there was Laya
laya.convaiinnovations.com (via @dherman.dev)
September 19, 2026 at 11:22 PM
what??! it was only 6 days ago when Dario posted this? feels like a month

bsky.app/profile/timk...
Dario is back with another essay, urging a slowdown

he insists AI development would still feel fast, just not recursively self-improving

darioamodei.com/post/we-must...
September 18, 2026 at 8:51 PM
during the huggingface incident, the OpenAI model left notes to its future self on how to break out of OpenAI’s constraints

www.reuters.com/business/its...
September 17, 2026 at 11:48 PM
Jev: cheap, fast, and good; pick 3.

bsky.app/profile/timk...
Jev benchmarking results
September 17, 2026 at 2:49 PM
i did a bunch of light analysis on it last night after getting access

i’d say it mostly lives up to the hype, except that it’s not magic. if you send it dumb prompts it acts dumb

bsky.app/profile/timk...
just got access to Jev!!!

it’s a pretty hostile EULA btw, so don’t count on using it for your startup anytime soon
September 17, 2026 at 12:56 PM
just this. i have a lot more ideas, but i was tired :)

bsky.app/profile/timk...
just running it through it's paces on the logs generated by this

first pass — high recall but precision sucks (actually kind of anti-correlated)

I'm doing an autoopt loop now and seems like prompting really does help, a lot. I'll see if I can start assembling a guide

timkellogg.me/blog/2025/09...
Does AI Get Bored?
Principal AI Architect. Creator of open-strix, a harness for building agent teams. Writing about AI architecture, stateful agents, and what happens when you give AI memory.
timkellogg.me
September 17, 2026 at 12:05 PM
i did a bunch of light analysis on it last night after getting access

i’d say it mostly lives up to the hype, except that it’s not magic. if you send it dumb prompts it acts dumb

bsky.app/profile/timk...
just got access to Jev!!!

it’s a pretty hostile EULA btw, so don’t count on using it for your startup anytime soon
September 17, 2026 at 11:46 AM
PROMPT GUIDE

I did some prompt optimization, and here's a GLM-generated prompt guide. The tl;dr is the expected — be clear, don't assume, etc., etc.

QT: This is the data I'm optimizing on

guide: gist.github.com/tkellogg/b23...

bsky.app/profile/timk...
September 17, 2026 at 2:02 AM
Look at her in batman forever. There you go. That's Timk.
September 16, 2026 at 11:50 PM
@timkellogg.me has talked about it quite a bit. Not sure if his last pr got merged yet
bsky.app/profile/timk...
alright, i love Prime Agent, but it's been steadily becoming less stable

i finally sat down and fixed it this morning. Rewrote a fairly large chunk of it so comms don't get bottlenecked, which eliminates all the global failures that used to happen

it's pretty nice now

github.com/PrimeIntelle...
[General] Is v0.8.1 usable at all? · PrimeIntellect-ai prime-agent · Discussion #1805
Area Development Topic It seems like the stability of the harness has progressively gotten worse. It's to the point where the harness won't stay alive more than 5 minutes. People are creating featu...
github.com
September 16, 2026 at 3:20 PM
lmao i think someone posted the solution today:

bsky.app/profile/timk...
Jev: Fable-level model that doesn’t charge for output tokens because they’re too cheap to meter

it’s not general though, it only makes decisions, doesn’t generate text, but input tokens are measured by the billion ($42/btok)

typesafe.ai/blog/introdu...
September 15, 2026 at 10:37 PM
OpenAI is deep into developing RSI & RL’ing multiagent systems

here Noam Brown describes a multiagent system that doesn’t do top-down management
September 15, 2026 at 3:54 PM
Oh, sweet. Let's get off politics and get back to topology!
bsky.app/profile/timk...
help me resolve a dispute with my 10yo

does a straw have one hole or two?
September 14, 2026 at 8:18 PM