Robust recipes to align language models with human and AI preferences
Multimodal Agents as Smartphone Users, an LLM-based multimodal agent
Open Source Deep Research Alternative to Reason and Search
Accelerate local LLM inference and finetuning
Tutorial tailored for Chinese babies on rapid fine-tuning
Pretrained time-series foundation model developed by Google Research
No-code LLM Platform to launch APIs and ETL Pipelines
In-depth tutorials on LLMs, RAGs and real-world AI agent applications
Ling-V2 is a MoE LLM provided and open-sourced by InclusionAI
Our first fully AI generated deep learning system
Large Audio Language Model built for natural interactions
95% token savings. 155x faster queries. 16 languages
Chinese XLNet pre-trained model
Inference script for Oasis 500M
Framework for building neural networks
StreamSpeech is a seamless model for offline speech recognition
Fast forecasting with statistical and econometric models
Generate Any 3D Scene in Seconds
Advanced techniques for RAG systems
Omnilingual ASR Open-Source Multilingual SpeechRecognition
A secure sandbox environment for malware developers and red teamers
A Model Context Protocol server for searching and analyzing arXiv
4M: Massively Multimodal Masked Modeling
This repository contains the official implementation of FastVLM
Refer and Ground Anything Anywhere at Any Granularity