+
+

Related Products

  • Vertex AI
    713 Ratings
    Visit Website
  • LM-Kit.NET
    17 Ratings
    Visit Website
  • Google AI Studio
    4 Ratings
    Visit Website
  • Amazon Bedrock
    72 Ratings
    Visit Website
  • Stack AI
    16 Ratings
    Visit Website
  • CDK Global
    331 Ratings
    Visit Website
  • Cody
    87 Ratings
    Visit Website
  • Nexo
    16,077 Ratings
    Visit Website
  • QuickApps
    Visit Website
  • Kinde
    48 Ratings
    Visit Website

About

This repository contains the research preview of LongLLaMA, a large language model capable of handling long contexts of 256k tokens or even more. LongLLaMA is built upon the foundation of OpenLLaMA and fine-tuned using the Focused Transformer (FoT) method. LongLLaMA code is built upon the foundation of Code Llama. We release a smaller 3B base variant (not instruction tuned) of the LongLLaMA model on a permissive license (Apache 2.0) and inference code supporting longer contexts on hugging face. Our model weights can serve as the drop-in replacement of LLaMA in existing implementations (for short context up to 2048 tokens). Additionally, we provide evaluation results and comparisons against the original OpenLLaMA models.

About

XLNet is a new unsupervised language representation learning method based on a novel generalized permutation language modeling objective. Additionally, XLNet employs Transformer-XL as the backbone model, exhibiting excellent performance for language tasks involving long context. Overall, XLNet achieves state-of-the-art (SOTA) results on various downstream language tasks including question answering, natural language inference, sentiment analysis, and document ranking.

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Platforms Supported

Windows
Mac
Linux
Cloud
On-Premises
iPhone
iPad
Android
Chromebook

Audience

Users interested in a powerful Large Language Model solution

Audience

Developers interested in a solution for generalized autoregressive pretraining for language understanding

Support

Phone Support
24/7 Live Support
Online

Support

Phone Support
24/7 Live Support
Online

API

Offers API

API

Offers API

Screenshots and Videos

Screenshots and Videos

Pricing

Free
Free Version
Free Trial

Pricing

Free
Free Version
Free Trial

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Reviews/Ratings

Overall 0.0 / 5
ease 0.0 / 5
features 0.0 / 5
design 0.0 / 5
support 0.0 / 5

This software hasn't been reviewed yet. Be the first to provide a review:

Review this Software

Training

Documentation
Webinars
Live Online
In Person

Training

Documentation
Webinars
Live Online
In Person

Company Information

LongLLaMA
github.com/CStanKonrad/long_llama

Company Information

XLNet
Founded: 2019
github.com/zihangdai/xlnet

Alternatives

Llama 2

Llama 2

Meta

Alternatives

BERT

BERT

Google
Mistral NeMo

Mistral NeMo

Mistral AI
GPT-4

GPT-4

OpenAI
DeepSeek-V2

DeepSeek-V2

DeepSeek

Categories

Categories

Integrations

Spark NLP

Integrations

Spark NLP
Claim LongLLaMA and update features and information
Claim LongLLaMA and update features and information
Claim XLNet and update features and information
Claim XLNet and update features and information