Concatenate a directory full of files into a single prompt
Qwen3-omni is a natively end-to-end, omni-modal LLM
Swirl queries any number of data sources with APIs
Let Claude (or any LLM) actually watch a video
A Telegram bot for Large Language Models
Big Model Application Development Practice 1
Running a 28.9M parameter LLM on an $8 microcontroller
Redundancy-aware KV Cache Compression for Reasoning Models
AI-driven multi-agent research assistant automating hypothesis
Mastering Applied AI, One Concept at a Time
LLM training in simple, raw C/CUDA
AI-powered code assistant for Vim. OpenAI and ChatGPT plugin for Vim
Linkedin Automation Tool
MemoryOS is designed to provide a memory operating system
Real-time multi-AI collaboration: Claude, Codex & Gemini
Recipes to train reward model for RLHF
Bringing BERT into modernity via both architecture changes and scaling
CV, NLP, LLM project applications, and advanced engineering deployment
95% token savings. 155x faster queries. 16 languages
Advanced techniques for RAG systems
A course of learning LLM inference serving on Apple Silicon
Tools for merging pretrained large language models
Toolkit for conversational AI
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
Run LLM prompts from your shell