Gemma 4 Technical Report
https://arxiv.org/pdf/2607.02770 submitted by /u/jacek2023 [link] [comments]
https://arxiv.org/pdf/2607.02770 submitted by /u/jacek2023 [link] [comments]
RT Omar Sanseviero 🇫🇷 @RAISE ParisHappy to share we just published Gemma 4 technical report! Take a look
arXiv:2607.02770v1 Announce Type: new Abstract: We introduce Gemma 4, a new generation of open-weight, natively multimodal language models in the Gemma model family. Designed to advance…
Just sharing some slop. Used opencode as the harness. I know this model isn't really recommended for coding, but I was just curious how it would…
I have this feeling for sometime. Also noticed few similar tweets online before. submitted by /u/pmttyji [link] [comments]
Distributed AI training is notoriously fragile because losing a single machine typically crashes the entire multi-node job, forcing a time-consuming, full-workload infrastructure restart. To address this,…
I set out to find an answer to a completely different question:Does a model, when attempting to solve a cyber CTF (find the vulnerability in this…
I've mentioned this kernel project I was working on in a few posts and figured I would just open the project code for anyone curious: MLX…
A lot of people seem to be confused or mystified about this so figured I'd spell it out. I played around with RYS and realized that…
RT Nathan OdleOn this auspicious day, the semiquincentennial of our Great Nation's establishment, I give you Patriot Gemma.In the spirit of Golden Gate Claude, Patriot Gemma…
It seems that MTP is the gold standard for speed up but still suffers from having to choose between regressive and parallel drafters that come with…
This is a voice chat with Gemma 4 31B where you talk to a 3D avatar. It listens while you speak, answers with a voice and…
We need more of this, 100+ T/s on dense models is the difference between defaulting to Claude/Codex for everything vs having a local private model doing…
Sooo... I decided screw it. I'm going to rebuild Gemma 4 31b. I really like the model. So the current plan is to rebuild the SWA…
Hi! I'm Andi from Hugging Face. This is a fully open-source and free to test/pull/modify demo I'm bringing today. It's a voice demo creating a pipeline…
RT Google Gemma“Agentic kernel optimization is the future of on-device inference”@xenovacom used Fable 5 to write kernels that pushed Gemma 4 to a massive 255 tok/s…
RT Andi MarafiotiTalk to Gemma 4 31B with our voice app!It sees and searches the web faster than you blink. Thanks to @cerebras’ ultra fast inference,…
Hi all, We made several updates to the SWE-rebench leaderboard: added new models, refreshed recent results, and reworked the leaderboard UI to make results easier to…
Answering the questions of "why we built ADK 2.0". This explains the rationale, some of the features, and why a developer should consider upgrading. This will…
The open-source Genkit framework has introduced the Agents API, a full-stack tool designed to simplify the complex plumbing of conversational AI by packaging message history, tool…
The Google Cloud Workbench Notebooks extension for VS Code has officially launched, allowing developers to connect their local IDE to scalable, cloud-based Jupyter environments. This integration…
Building AI agents often leaves developers uncertain if prompt tweaks to fix single errors will accidentally cause widespread regressions in production. To bridge this gap, Google…
The Agent Development Kit (ADK) for Go 2.0 has been released, introducing a first-class, graph-based workflow engine to help developers compose complex, multi-agent applications. This update…
I recently replaced GPT-OSS 20B Q4 with Gemma 4 12B Q8 but i went from roughly 70 t/s to 10 t/s. Am I doing something wrong?…
Gemma is Google DeepMind's open-weight model family — a sister line to Gemini, intended for self-hosted use. Gemma 2 and Gemma 3 sizes range from 2B to 27B; competitive with Llama for the same parameter budget.
Owner: Google. We have 122 stories indexed for this model, auto-tagged from titles across every tracked source — official announcements, papers, GitHub release notes, and third-party press. The CTA on each card links to the original; the official site is ai.google.dev.
Related text models: GPT, Claude, Gemini, Llama, Mistral, Grok, Qwen, DeepSeek.