Compare the Top Large Language Models that integrate with ExecuTorch as of August 2026

This a list of Large Language Models that integrate with ExecuTorch. Use the filters on the left to add additional filters for products that have integrations with ExecuTorch. View the products that work with ExecuTorch in the table below.

What are Large Language Models for ExecuTorch?

Large language models are artificial neural networks used to process and understand natural language. Commonly trained on large datasets, they can be used for a variety of tasks such as text generation, text classification, question answering, and machine translation. Over time, these models have continued to improve, allowing for better accuracy and greater performance on a variety of tasks. Compare and read user reviews of the best Large Language Models for ExecuTorch currently available using the table below. This list is updated regularly.

  • 1
    Muse Glimmer
    Muse Glimmer is a 30-billion-parameter open-weights model from Meta Superintelligence Labs, optimized for always-on local agent workflows. Small enough to run on a Mac or PC with a single consumer GPU, it is designed for local agents, function calling, coding, and LLM-as-a-judge evaluation without depending on cloud infrastructure or network access. The model combines long-horizon execution, precise tool calling, multimodal understanding, long-context memory, and instruction following. It can complete end-to-end agentic tasks, sustain multi-step reasoning across extended workflows, recover from failed or unexpected tool calls, and accept interleaved text and images through a dedicated perception encoder for interpreting screenshots, charts, and documents. Muse Glimmer works with OpenClaw and other agentic orchestration patterns, supports controllable reasoning effort, and is trained on data from more than 100 languages.
    Starting Price: Free
  • 2
    Llama 3.2
    The open-source AI model you can fine-tune, distill and deploy anywhere is now available in more versions. Choose from 1B, 3B, 11B or 90B, or continue building with Llama 3.1. Llama 3.2 is a collection of large language models (LLMs) pretrained and fine-tuned in 1B and 3B sizes that are multilingual text only, and 11B and 90B sizes that take both text and image inputs and output text. Develop highly performative and efficient applications from our latest release. Use our 1B or 3B models for on device applications such as summarizing a discussion from your phone or calling on-device tools like calendar. Use our 11B or 90B models for image use cases such as transforming an existing image into something new or getting more information from an image of your surroundings.
    Starting Price: Free
  • 3
    LLaVA

    LLaVA

    LLaVA

    LLaVA (Large Language-and-Vision Assistant) is an innovative multimodal model that integrates a vision encoder with the Vicuna language model to facilitate comprehensive visual and language understanding. Through end-to-end training, LLaVA exhibits impressive chat capabilities, emulating the multimodal functionalities of models like GPT-4. Notably, LLaVA-1.5 has achieved state-of-the-art performance across 11 benchmarks, utilizing publicly available data and completing training in approximately one day on a single 8-A100 node, surpassing methods that rely on billion-scale datasets. The development of LLaVA involved the creation of a multimodal instruction-following dataset, generated using language-only GPT-4. This dataset comprises 158,000 unique language-image instruction-following samples, including conversations, detailed descriptions, and complex reasoning tasks. This data has been instrumental in training LLaVA to perform a wide array of visual and language tasks effectively.
    Starting Price: Free
  • 4
    Qwen3

    Qwen3

    Alibaba

    Qwen3, the latest iteration of the Qwen family of large language models, introduces groundbreaking features that enhance performance across coding, math, and general capabilities. With models like the Qwen3-235B-A22B and Qwen3-30B-A3B, Qwen3 achieves impressive results compared to top-tier models, thanks to its hybrid thinking modes that allow users to control the balance between deep reasoning and quick responses. The platform supports 119 languages and dialects, making it an ideal choice for global applications. Its pre-training process, which uses 36 trillion tokens, enables robust performance, and advanced reinforcement learning (RL) techniques continue to refine its capabilities. Available on platforms like Hugging Face and ModelScope, Qwen3 offers a powerful tool for developers and researchers working in diverse fields.
    Starting Price: Free
  • Previous
  • You're on page 1
  • Next