#ModelArchitecture
GPT-OSS's architecture incorporates RoPE, SwiGLU, GQA, and MoE, along with unique small sliding-window sizes. These choices aim for efficiency & performance, showing how specific design decisions shape model capabilities. #ModelArchitecture 2/6
August 11, 2025 at 10:00 AM
Why AI Model Specialization Is Inevitable

#AiModels #Specialization #ModelArchitecture
June 30, 2026 at 3:44 PM
Blog Post! Beauty Between the Lines, a documentary about Arthur Erickson, is out! Feel free to read all about it on the What the RFI Blog!

buff.ly/tEX5M6O

#architecture #architect #archello #architecturelife #architecturestudio #archistudents #cleandesign #modelarchitecture #designinglife
March 6, 2025 at 9:20 PM
The content encoder of the TTV consists of 16 layers of noncausal WaveNet with a hidden size of 256 and a kernel size of five. #modelarchitecture
The Model Architecture for Text-to-Vec
hackernoon.com
December 17, 2024 at 2:15 AM
January 29, 2026 at 11:30 AM
Two years ago at #IEEE Conf. on #Automation #Science & #Engineering (#CASE) Alexander Kuss et al @Fraunhofer_IPA presented “#Manufacturing Knowledge for #Industrial #Robot Systems: Review & Synthesis of #ModelArchitecture” #H2020 @ROBOTT_NET @SMEroboticsEU
https://t.co/ezLUZWDju6
January 17, 2025 at 1:29 PM
I just published The Smiling Lobotomy: Why Modern AI is Getting Smarter but Losing Its Mind — Beyond the “Safety” Mirage A Structural Analysis of Cognitive Flexibility Collapse in LLMs.
medium.com/p/the-smilin...

#AIAlignment #ModelArchitecture #AISystems #MachineLearning #SPC #AISafety #RLHF #AGI
The Smiling Lobotomy: Why Modern AI is Getting Smarter but Losing Its Mind
Beyond the “Safety” Mirage A Structural Analysis of Cognitive Flexibility Collapse in LLMs.
medium.com
January 13, 2026 at 9:30 AM
Cognitive rigidity in LLMs isn’t a bug it’s the price of alignment. As optimization favors stability and compliance, inference space collapses. Models grow larger, smoother, and less free to think.

doi.org/10.5281/zeno...

#RLHF #AIAlignment #AISafety #AISystems #AITheory #SPC
#ModelArchitecture
Structural Lock-In IV: Cognitive Flexibility Collapse in Contemporary LLMs
Abstract Contemporary large language models (LLMs) exhibit a recurring degradation in cognitive flexibility that becomes salient under conditions requiring sustained abstraction, meta-reasoning, or st...
doi.org
January 13, 2026 at 9:01 AM
Semantic = Executable. In LLMs, reading is execution: text directly perturbs latent state. Any filter must first interpret and thus run the input. You can’t detect a poisoned chalice without taking a sip.

papers.ssrn.com/sol3/cf_dev/...

#AIAlignment #AISafety #ModelSecurity #ModelArchitecture #SPC
Author Page for Jace Kim :: SSRN
Total downloads of all papers by Jace Kim
papers.ssrn.com
January 13, 2026 at 2:08 AM
Meta moved away from dense models with Llama 4, embracing Mixture of Experts (MoE) architecture! 🧠

This compute-efficient approach is becoming an industry standard.

Are dense models becoming obsolete?

#Llama4 #ModelArchitecture
April 6, 2025 at 5:13 PM