HunyuanOCR
Tencent Hunyuan is a large-scale, multimodal AI model family developed by Tencent that spans text, image, video, and 3D modalities, designed for general-purpose AI tasks like content generation, visual reasoning, and business automation. Its model lineup includes variants optimized for natural language understanding, multimodal vision-language comprehension (e.g., image & video understanding), text-to-image creation, video generation, and 3D content generation. Hunyuan models leverage a mixture-of-experts architecture and other innovations (like hybrid “mamba-transformer” designs) to deliver strong performance on reasoning, long-context understanding, cross-modal tasks, and efficient inference. For example, the vision-language model Hunyuan-Vision-1.5 supports “thinking-on-image”, enabling deep multimodal understanding and reasoning on images, video frames, diagrams, or spatial data.
Learn more
Kimi K3
Kimi K3 is Moonshot AI’s most capable model, built for frontier intelligence scenarios such as software engineering, knowledge work, deep reasoning, and multimodal understanding. The model has 2.8 trillion parameters and uses Kimi Delta Attention, a hybrid linear attention mechanism, along with Attention Residuals for long-context performance. Kimi K3 supports a 1 million token context window, making it useful for analyzing large codebases, long documents, complex knowledge bases, and multi-step workflows. It includes native visual understanding for images and videos, with support for structured message formats, base64 image input, uploaded video files, and multimodal reasoning. Developers can use Kimi K3 through an OpenAI-compatible API with support for streaming, structured JSON output, partial mode, custom tools, dynamic tool loading, and automatic context caching.
Learn more
LiveKit
LiveKit is a real-time platform that enables developers to build video, voice, and data capabilities into their applications. Building on WebRTC, it supports a broad range of frontend and backend platforms. LiveKit's network is optimized for ultra-low latency, extreme resiliency, and massive scale. Our team is distributed across the world, and our infrastructure delivers billions of minutes of audio and video every month. LiveKit provides SDK support across all major platforms, allowing you to code your application with a LiveKit client natively designed for your platform of choice. You can self-host LiveKit for free without changing a line of code, as the entire ecosystem of tools and services is Apache 2.0 open source. LiveKit offers a feature-rich platform, including SSO and RBAC for teams, enterprise-grade security with end-to-end encryption, noise and echo cancellation, session recording, stream ingestion, and moderation tools.
Learn more
InfiEye
AI-video analytics enables your store managers to identify and prevent store shrinkages as well as inventory thefts, as soon as they happen. With InfiEye AI, you can improve your in-store shopping experience by identifying fast-selling SKUs on retail shelves and easily monitor in-store customer behaviors. Integrable with your existing in-store PoE cameras. Place your cameras at points you want to monitor, be it the overhead of checkout counters, countertop, on retail shelves, or at the store's entry and exit points. Image recognition algorithm processes live in-store feeds, frame by frame, to accurately identify every object on the shop floor. Evidence-based alerts are sent to the store staff to intervene in a friendly manner. Monitor inventory stock-outs or overstocking, and track sales performance of each store. Reduce store shrinkages and improve net-sales output of every store.
Learn more