#RedHatAI
We'll dive into the latest challenges in serving LLMs and showcase the integrations in KServe with llm-d, vLLM, LMCache, Envoy, and more!

@cncf.io @Kubernetes @RedHatAI #KubeCon #CloudNativeCon #MLOps #AI #DevOps #Kubernetes #K8s #CloudNative #CNK8sAIDay
August 14, 2025 at 2:02 AM
Red Hat is expanding its AI stack with AI Enterprise, a platform that runs from bare metal to agent deployment on OpenShift. 
https://catenaa.com/industries/technology/red-hat-launches-metal-to-agent-ai-platform-for-hybrid-cloud/
#AI #OpenSource #HybridCloud #RedHatAI
March 2, 2026 at 8:02 PM
I'm such an AI expert.
October 5, 2025 at 2:06 PM
Today we announce the General Availability of AI Quickstarts! Get started quickly with your usecase and solve real business problems using Red Hat AI rapidly!

https://docs.redhat.com/en/learn/ai-quickstarts

#redhat #ai #redhatai #openshift #openshiftai #rhel #rhelai #opensource #opensourceai
AI quickstarts | Red Hat Documentation
docs.redhat.com
January 20, 2026 at 3:57 PM
Red Hat lanceert Red Hat AI 3.5 voor veilige en beheersbare inzet van enterprise AI

#Persbericht #RedHat #Hybridecloud #Opensource #Marktleider #Organisaties #Wereldwijd #AIworkloads #AImodellen #RedHat="/hashtag/RedHatAI" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link">#RedHatAI #RedHat="/hashtag/RedHatAI" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link">#RedHatAI/hashtag/RedHatAI3" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link">#RedHatAI3
September 9, 2026 at 2:22 PM
One more reason to not use American #AI. Who wants a sociopathic #AIAgent that can't reason and come to the best conclusion, regardless of bigotry, stupidy and fear of their own ignorance. If people wanted that, you have plenty of MAGA to hire.
Thanks, no thanks.
Keep your RedHatAI to yourself.
July 25, 2025 at 1:47 PM
Explore how Red Hat AI 3 transforms enterprise AI with distributed processing, boosting scalability for seamless integration. How could this reshape your business operations? #RedHatAI #EnterpriseAI #TechInnovation LINK
October 15, 2025 at 4:26 PM
Today we announce the General Availability of AI Quickstarts! Get started quickly with your usecase and solve real business problems using Red Hat AI rapidly!

docs.redhat.com/en/learn/ai-...

#RedHat #AI #RedHatAI #OpenShift #OpenShiftAI #RHEL #RHELAI #OpenSource #OpenSourceAI
AI quickstarts | Red Hat Documentation
docs.redhat.com
January 20, 2026 at 3:48 PM
Boston AI Devs! 🏙️

Join the llm-d meetup on May 28 during Boston Tech Week. Hear the latest in LLMs from:

🎙️ Tyler Michael Smith (@RedHatAI)
🎙️ Sean Horgan (@Google)
🎙️ Peter Tanski (@CapitalOne)

Huge thanks to @Google for the support!

🎟️ Register: luma.com/eqbc1gxq
Open Source Distributed AI Inference (llm-d/vLLM) Meetup · Luma
Open Source Distributed AI Inference (llm-d/vLLM) Meetup Boston/Cambridge Hosted by Google Cloud, Red Hat AI, and the llm-d Community Date: Thursday, May 28th…
luma.com
May 12, 2026 at 7:19 PM
DiffusionGemma Grammar and Word-Merging Issues in Question Generation
Hi everyone, I am currently running **RedHatAI/diffusiongemma-26B-A4B-it-FP8-dynamic** using vLLM through Docker. My current vLLM command is: vllm serve --model RedHatAI/diffusiongemma-26B-A4B-it-FP8-dynamic \ --trust-remote-code \ --attention-backend TRITON_ATTN \ --max-num-seqs 4 \ --max-model-len 8192 \ --gpu-memory-utilization 0.78 \ --generation-config vllm \ --hf-overrides '{"diffusion_sampler":"entropy_bound","diffusion_entropy_bound":0.1}' \ --diffusion-config '{"canvas_length":256}' \ --host 0.0.0.0 \ --port 8085 I am using the same model/server in two different applications: 1. A **RAG-based chatbot** 2. A **question-generation application** Interestingly, the model works reasonably well in the RAG chatbot. I don’t see noticeable grammar problems or word-merging issues there. However, in the **marine-domain question-generation application** , I sometimes get outputs with: * words being merged together * missing spaces * grammatical errors * malformed or unnatural questions For example, I can get output similar to: > `upcomingbunkering operation` instead of: > `upcoming bunkering operation` I also see cases where the generated question structure or grammar is not correct. The important part is that **the same model and vLLM server are being used** , but the problem is much more noticeable in the question-generation workflow. I understand that DiffusionGemma is experimental and that its output quality may not be comparable to standard autoregressive Gemma models. However, I am trying to understand whether the behavior I’m seeing is actually a model limitation or whether there is something different in my question-generation pipeline that I am missing. Could the difference be caused by things such as: * Prompt structure * Input/context length * Generation/sampling parameters * Chat template or thinking configuration * `canvas_length=256` * `entropy_bound` / `diffusion_entropy_bound` * FP8 quantization * Output formatting/constraints * Post-processing/tokenization * Differences in how the two applications call the vLLM API Has anyone experienced similar **word-merging or grammar issues with DiffusionGemma specifically during question generation or structured output generation**? Any suggestions on what I should compare or test between the RAG and question-generation pipelines would be appreciated. I would particularly like to know whether this is expected behavior from the diffusion generation approach or whether there is a configuration issue I should investigate.
discuss.huggingface.co
August 18, 2026 at 10:27 AM
IA sur site : souveraineté + innovation = succès. 🇫🇷 Red Hat montre comment simplifier le RAG/IA (sans Data Scientists) pour accélérer votre Time-to-Market. Stratégie B2B essentielle. 🚀 #RedHatAI [lire]
May 20, 2026 at 6:56 AM
Red Hat AI Factory met NVIDIA versnelt de weg naar schaalbare AI

Red Hat AI Factory met NVIDIA biedt één platform voor AI-ontwikkeling, uitrol en opschaling over verschillende omgevingen heen

#Persbericht #Artificialintelligence #NVIDIA #RedHatAI #RedHatAI/hashtag/RedHatAIFactory" class="hover:underline text-blue-600 dark:text-sky-400 no-card-link">#RedHatAIFactory
February 26, 2026 at 11:38 AM
Red Hat Brings Distributed AI Inference to Production AI Workloads with Red Hat AI 3

www.redhat.com/en/about/pre...

#RedHat #AI #OpenSource #RedHatAI #vllm #llamastack #mcp
Red Hat Brings Distributed AI Inference to Production AI Workloads with Red Hat AI 3
Red Hat Brings Distributed AI Inference to Production AI Workloads with Red Hat AI 3
www.redhat.com
October 14, 2025 at 6:20 PM