Skip to content

RightNow Research Lab

Enabling Model-Hardware Co-Design at Scale

RightNow AI

Abstract

We build tools that close the gap between AI models and the hardware they run on.

NVIDIA Inception member · 6 papers on arXiv

NVIDIA
AMD
Meta
Hugging Face
Google
Runway
Endorsement
Director of Accelerated Computing · NVIDIA
Figure 1: NVIDIA's Director of Accelerated Computing, on AI-generated CUDA kernels.

Products

Recorded on runinfra.ai on 2026-10-09: the home page, Open models, built for agents, with a model's speed, price and cache-hit panel, then that model's API call in Python and the Anthropic SDK.

Hosted open models

RunInfra

One API key for the OpenAI and Anthropic SDKs. Billed per token from your balance.

Explore RunInfra (opens in a new tab)
Figure 2: runinfra.ai, recorded 2026-10-09.

Research

Figure 3: The compute problem. Illustrative GPU utilization before and after optimized kernels.
Agent OS

OpenFang

Rust

Low-level agent execution with direct OS and GPU access.

RightNow-AI/openfang
OpenFang: GitHub Stars
18,209
GitHub API count: 2026-09-21; chart illustrative
rightnow
RightNow Logo

Join the community

Enabling Model-Hardware Co-Design at Scale

JOIN DISCORD