Ongoing research training transformer models at scale
Tools like web browser, computer access and code runner for LLMs
Test-Time Reinforcement Learning
Minimal reproduction of OneRec
AI-powered penetration testing assistant using local LLM on linux
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
Using AI models to automatically provide commentary and edit videos
The Multi-Agent Framework
Framework and no-code GUI for fine-tuning LLMs
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
On the Structural Pruning of Large Language Models
A.S.E (AICGSecEval) is a repository-level AI-generated code security
Towards Efficient Self-Evolving Agent System
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
Scalable RL solution for advanced reasoning of language models
Gemma open-weight LLM library, from Google DeepMind
Benchmark LLMs by fighting in Street Fighter 3
Production-grade platform for building agentic IM bots
One-stop solution for creating your digital avatar from chat history
Recipes to train reward model for RLHF
Bringing BERT into modernity via both architecture changes and scaling
CV, NLP, LLM project applications, and advanced engineering deployment
SimpleMem: Efficient Lifelong Memory for LLM Agents