Advanced AI Explainability for computer vision
Transfer learning / domain adaptation / domain generalization
Learn how to develop, deploy and iterate on production-grade ML
AI Agent Evaluator & Red Team Platform
Agent framework that enables tool-use agent tasks
Convert websites into structured APIs automatically with Python tool
Extension of Google Research’s PaperBanana
A large-scale model of medical consultation in Chinese
LongBench v2 and LongBench (ACL 25'&24')
Uncertainty Quantification for Language Models, is a Python package
Streamlines and simplifies prompt design for both developers
Linkedin Automation Tool
Hypernetworks that adapt LLMs for specific benchmark tasks
MemoryOS is designed to provide a memory operating system
Unified KV Cache Compression Methods for Auto-Regressive Models
Learning to Reason with Search for LLMs via Reinforcement Learning
OSINT tool for discovering email addresses in known data breaches
Benchmark LLMs by fighting in Street Fighter 3
Cache-Augmented Generation: A Simple, Efficient Alternative to RAG
270+ Claude Code plugins with 739 agent skills
Recipes to train reward model for RLHF
A tension reasoning engine over 131 S-class problems
Bringing BERT into modernity via both architecture changes and scaling
Scalable RL solution for advanced reasoning of language models
An agentless approach to automatically solve software development