Efficient Triton Kernels for LLM Training
- Updated
Jul 23, 2026 - Python
Efficient Triton Kernels for LLM Training
Explore LLM model deployment based on AXera's AI chips
Gemma2(9B), Llama3-8B-Finetune-and-RAG, code base for sample, implemented in Kaggle platform
RAG-based Telegram assistant bot for freshmen
Boost RAG performance with question decomposer
Gemma2 2B model that fine tuned with an e-commerce data.
This project focuses on efficient machine translation for nine Indic languages using the fine-tuned Gemma2-2B LLM and adapter switching, reducing computational overhead. It also leverages agentic methods and the Groq API for quality assurance and accurate translation analysis between source and target segments.
A complete guide to NLP and ML for text processing, covering rule-based models, RNNs, CNNs, Transformers, entity detection, sentiment analysis, LLM fine-tuning, RAG, and prompt engineering with tools like Langchain and Ollama.
AI Discord Bot (GEMM-X) is an intelligent assistant for Discord, leveraging AI technologies from multiple providers to generate images, create music, produce speech, and more. It supports custom personality settings and advanced user/server configurations.
A chatbot created using huggingface's gemma 2 created by Google. Authentication using Supabase Authentication.