LLaMA-Factory

LLaMA-Factory

hoshi-hiyouga
Ludwig

Ludwig

Uber AI
+
+

Related Products

  • Pipedrive
    10,536 Ratings
    Visit Website
  • Checksum.ai
    1 Rating
    Visit Website
  • LTX
    182 Ratings
    Visit Website
  • Runpod
    230 Ratings
    Visit Website
  • Google AI Studio
    30 Ratings
    Visit Website
  • Gemini Enterprise Agent Platform
    999 Ratings
    Visit Website
  • LM-Kit.NET
    29 Ratings
    Visit Website
  • KrakenD
    71 Ratings
    Visit Website
  • StackAI
    54 Ratings
    Visit Website
  • AnalyticsCreator
    46 Ratings
    Visit Website

About

​LLaMA-Factory is an open source platform designed to streamline and enhance the fine-tuning process of over 100 Large Language Models (LLMs) and Vision-Language Models (VLMs). It supports various fine-tuning techniques, including Low-Rank Adaptation (LoRA), Quantized LoRA (QLoRA), and Prefix-Tuning, allowing users to customize models efficiently. It has demonstrated significant performance improvements; for instance, its LoRA tuning offers up to 3.7 times faster training speeds with better Rouge scores on advertising text generation tasks compared to traditional methods. LLaMA-Factory's architecture is designed for flexibility, supporting a wide range of model architectures and configurations. Users can easily integrate their datasets and utilize the platform's tools to achieve optimized fine-tuning results. Detailed documentation and diverse examples are provided to assist users in navigating the fine-tuning process effectively.

About

Ludwig is a low-code framework for building custom AI models like LLMs and other deep neural networks. Build custom models with ease: a declarative YAML configuration file is all you need to train a state-of-the-art LLM on your data. Support for multi-task and multi-modality learning. Comprehensive config validation detects invalid parameter combinations and prevents runtime failures. Optimized for scale and efficiency: automatic batch size selection, distributed training (DDP, DeepSpeed), parameter efficient fine-tuning (PEFT), 4-bit quantization (QLoRA), and larger-than-memory datasets. Expert level control: retain full control of your models down to the activation functions. Support for hyperparameter optimization, explainability, and rich metric visualizations. Modular and extensible: experiment with different model architectures, tasks, features, and modalities with just a few parameter changes in the config. Think building blocks for deep learning.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

AI researchers and developers wanting a solution to fine-tune a wide array of language and vision-language models

Audience

Developers interested in a low-code framework to build custom AI models like LLMs and other deep neural networks

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

No information available.
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

hoshi-hiyouga
github.com/hiyouga/LLaMA-Factory

Company Information

Uber AI
Founded: 2016
United States
ludwig.ai/latest/

Alternatives

Alternatives

DeepSpeed

DeepSpeed

Microsoft
MLBox

MLBox

Axel ARONIO DE ROMBLAY
Lens

Lens

Moondream

Categories

Categories

Integrations

MLflow
TensorBoard
Aim
Alpaca
Comet
DeepSeek
Docker
Gemma
Hugging Face
Llama
Llama 2
Mistral AI
Mixtral 8x7B
Phi-2
Python
Qwen
RAY
TensorWave
Weights & Biases
Yi-Large

Integrations

MLflow
TensorBoard
Aim
Alpaca
Comet
DeepSeek
Docker
Gemma
Hugging Face
Llama
Llama 2
Mistral AI
Mixtral 8x7B
Phi-2
Python
Qwen
RAY
TensorWave
Weights & Biases
Yi-Large
Claim LLaMA-Factory and update features and information
Claim LLaMA-Factory and update features and information
Claim Ludwig and update features and information
Claim Ludwig and update features and information