Transfer learning / domain adaptation / domain generalization
Learn how to develop, deploy and iterate on production-grade ML
AI Agent Evaluator & Red Team Platform
Agent framework that enables tool-use agent tasks
Convert websites into structured APIs automatically with Python tool
Extension of Google Research’s PaperBanana
A large-scale model of medical consultation in Chinese
LongBench v2 and LongBench (ACL 25'&24')
Uncertainty Quantification for Language Models, is a Python package
Streamlines and simplifies prompt design for both developers
Linkedin Automation Tool
Hypernetworks that adapt LLMs for specific benchmark tasks
MemoryOS is designed to provide a memory operating system
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
OSINT tool for discovering email addresses in known data breaches
Benchmark LLMs by fighting in Street Fighter 3
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
270+ Claude Code plugins with 739 agent skills
Recipes to train reward model for RLHF
A tension reasoning engine over 131 S-class problems
Bringing BERT into modernity via both architecture changes and scaling
Scalable RL solution for advanced reasoning of language models
An agentless approach to automatically solve software development
Neural Network architecture based on ideas of the original LSTM