
4/9/2024 · Tris Warkentin, Jane Fine
What this post added
Introduces CodeGemma, a 7B pretrained and instruction-tuned variant, and a 2B pretrained variant, specialized for code completion and generation tasks. Also introduces RecurrentGemma, an efficiency-optimized architecture leveraging recurrent neural networks and local attention for improved memory efficiency and higher throughput, showcasing a non-transformer model. Updates Gemma 1.1 with performance improvements and bug fixes. Details compatibility with various frameworks and hardware.