Laya-MLX is an independent MLX implementation of Laya’s typed decision models for Apple Silicon Macs. It performs structured choice, score, and probability decisions without token-by-token text generation. Inference runs fully locally after model weights are downloaded and does not require PyTorch, Transformers, or a cloud API. The runtime supports English, multilingual, and typed-decision Laya checkpoints while preserving their original calibration and output formats. Its implementation moves the encoder, decision transformer, scoring head, and action head into MLX. The project reports short-decision latency in the single-digit to low-teens millisecond range on an M3 Max, depending on the checkpoint. It also includes routing utilities, demos, validation tests, benchmarks, and a Python API.

Features

  • Native Apple Silicon MLX inference
  • Choice, score, and probability decisions
  • Fully local execution
  • English and multilingual checkpoints
  • Python API and model routing
  • Benchmarks, validation tests, and demos

Project Samples

Project Activity

See All Activity >

Categories

AI Models

License

Apache License V2.0

Follow Laya-MLX

Laya-MLX Web Site

Other Useful Business Software
MongoDB Atlas runs apps anywhere Icon
MongoDB Atlas runs apps anywhere

Deploy in 115+ regions with the modern database for every enterprise.

MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
Start Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of Laya-MLX!

Additional Project Details

Programming Language

Python

Related Categories

Python AI Models

Registered

13 hours ago