Test-Time Reinforcement Learning
Minimal reproduction of OneRec
Using AI models to automatically provide commentary and edit videos
Framework and no-code GUI for fine-tuning LLMs
AI-powered penetration testing assistant using local LLM on linux
Benchmark LLMs by fighting in Street Fighter 3
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
Gemma open-weight LLM library, from Google DeepMind
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
On the Structural Pruning of Large Language Models
A.S.E (AICGSecEval) is a repository-level AI-generated code security
Towards Efficient Self-Evolving Agent System
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
Scalable RL solution for advanced reasoning of language models
MobileLLM Optimizing Sub-billion Parameter Language Models
Production-grade platform for building agentic IM bots
One-stop solution for creating your digital avatar from chat history
A system for agentic LLM-powered data processing and ETL
Universal LLM Deployment Engine with ML Compilation
Recipes to train reward model for RLHF
Bringing BERT into modernity via both architecture changes and scaling
CV, NLP, LLM project applications, and advanced engineering deployment
SimpleMem: Efficient Lifelong Memory for LLM Agents