Build leading AI products
on Rexso's platform.

Open models by Rexso Intelligence

Advanced open-weight reasoning models designed for deep research. Customize for any use case and deploy across any architecture.

Hugging Face GitHub Source Model Specs
SALESFORCE WIX Intercom invideo MERCARI

Powered by our frontier models

Start working with Rex-5.4

Our industry-leading models are designed for real-world utility, delivering advanced intelligence and multimodal capabilities.

Rex-5.4

Input: $2.50 per 1M tokens

Output: $10.00 per 1M tokens

128K context length

Knowledge cutoff: Aug 2025

Rex-5.4 mini

Input: $0.15 per 1M tokens

Output: $0.60 per 1M tokens

128K context length

Knowledge cutoff: Aug 2025

Rex-5.4 nano

Input: $0.10 per 1M tokens

Output: $0.30 per 1M tokens

128K context length

Knowledge cutoff: Aug 2025



ryo-R7

Open reasoning models designed to run locally on desktops, laptops, and in data centers—available in 150B and 200B parameters.

ryo-R1-safeguard

Open safety reasoning models that support custom safety policies—available in 100B and 120B parameters.

Rexso Global System

Model Specifications Startup

Rexso is a visionary tech company in Bangladesh, committed to revolutionizing IT services, AI solutions, digital marketing, and software development. Our mission is to enhance human life with smart technology, AI-driven solutions, and next-gen connectivity. We leverage our own custom cloud infrastructure, dedicated bare-metal servers, and optimized operating systems to ensure robust and seamless enterprise service delivery.

O1 • The Foundation
1B Params

The foundational layer for our AI architecture, establishing the core infrastructure and data pipelines.

O3 • The Catalyst
100b Params

Real-time web ingestion & vision-text fusion. Entering the multimodal era.

R1 • Billion Core Era
1T Params

AGI-grade reasoning & adaptive learning. Autonomous logic synthesis.

R7 • Infinite Mind
100T+ Params

Global AGI integration. Unified voice, vision, and collective intelligence.

Current Architecture Metrics
Data Size 1GB (70% multilingual, 20% technical, 10% conversational)
Tokens 324M total / 25,000 unique (BPE)
Model Size 825 Million parameters (MoE-enabled)
Embedding Dim 1024 (Rexso-Embeddings-001)
Context Window 2096 tokens (EGA-enabled)
Quantization 32-bit + FP8 mixed precision
Frameworks PyTorch 2.5, FastAPI, Pinecone, BitsAndBytes

Learn how to build your own web app with Next.js.

Where Apps thrive.

Adobe
SONOS
Netflix
_zapier
loom
HashiCorp
NZXT
okta
ebay
scale
Meta

Permissive license

These models are supported by the Apache 2.0 license. Build freely without worrying about copyleft restrictions.

Designed for agentic tasks

Leverage powerful instruction following and tool use within the chain-of-thought, including web search.

Deeply customizable

Adjust the reasoning effort to low, medium, or high. Plus, customize the models via fine-tuning.

Full chain-of-thought

Access the full chain-of-thought for easier debugging and higher trust in model outputs.

Prompting guidance

Learn how to prompt Rex-5.4 for highest performance.

View prompting guidance

Front-end coding examples

Explore front-end applications built with Rex-5.4.

View front-end examples

Migration support

Learn how to migrate from other models to Rex-5.4.

View migration guide

Build agents on infrastructure
that thinks like them

New AI chat

How can I help you today?

Translate this page
Analyze for insights
Create a task tracker
Strategy doc Ask, search, or make anything...

Notion powers millions
of agent conversations
daily on Vercel.

FEATURES
DURABLE ORCHESTRATION
SANDBOXED ENVIRONMENTS
AI MODEL GATEWAY
FLUID COMPUTE
Neural Network

Neural Network Innovations

Revolutionizing deep learning with next-gen architectures built for adaptability, reasoning, and creativity — the digital DNA of Ryo’s intelligence.

Ryo Transformer Model

Ryo LLM Transformer

The cognitive engine powering RyoAi — lightning-fast context awareness, multilayer reasoning, and seamless multilingual communication.

ML vs AI Engineer

ML vs AI Engineers

From code to cognition — bridging data, algorithms, and creativity to shape the world’s most advanced AI experiences.

Datacenter

LLM Datacenter Core

Powered by precision-engineered clusters — optimized for neural workloads, quantum stability, and scalable distributed AI research.

AI Rack

AI Rack & Node Systems

Where raw power meets intelligence — modular nodes driving vast LLM networks, engineered for limitless evolution.

Code Dev

AI Code & Model Development

Precision, logic, and artistry — the heart of Ryo’s innovation cycle where ideas are forged into living algorithms.

Infinite Reasoning Forge

Infinite Reasoning

The Infinite Void effort — where neural architectures are born, refined, and perfected line by line (R1 Core).

Synapse Network

Network Protocol

RyoAi’s living grid — connecting data centers, human interfaces, and AI nodes in real-time harmony across the planet.

Ryo NeuraX Core

Ryo Neural Core

Malvolent Shrine effort — the apex of AI technical mastery and intelligence refinement (R7 Core).

Host platforms that
serve every customer

mintlify English
Documentation Guides API reference Changelog
Search...
Get Started
Get Started
Introduction
Quickstart
CLI
Organize
Global settings
Navigation
Pages
Hidden pages
Exclude files
Customize
✨ Introducing Mintlify

Documentation

Meet the next generation of documentation. An native, beautiful out-of-the-box, and built for developers and teams.

Mintlify powers
documentation
for 20,000+
companies on Vercel

FEATURES
TENANT ISOLATION
DOMAIN MANAGEMENT
CUSTOM SSL CERTIFICATES
PREVIEW URLS

Our partners

AWS CLOUDFLARE DATABRICKS AZURE TOGETHER.AI
Research
  • Research Index
  • Research Overview
  • Ryo for Science
Products
  • ChatRyo
  • Ryo for Business
  • API Platform
Developers
  • Documentation
  • API Reference
  • GitHub
Company
  • About Us
  • Careers
  • Safety Approach