creda/inference

An independent inference project

Open models.
A new place
to run them.

We’re building Creda Inference: a hosted API for open-weight models, with a focus on useful capacity and clear performance data.

In development. API access is not yet available.

THE SERVICE WE’RE BUILDINGConcept

From model selection to a dependable endpoint.
One focused launch at a time.

OUR APPROACH

Start focused. Measure what matters.

01

Useful model access

We’re choosing an initial open-weight model around real inference needs. The model catalog will be published when serving is ready.

02

A familiar interface

Our planned API follows the OpenAI-compatible format, with streaming responses and transparent token usage.

03

Evidence before promises

Measured latency, throughput, and pricing will accompany the launch. Our data-handling policy will be available before API access opens.

CURRENT STATUS

Building toward
the first endpoint.

Prelaunch

Model selection, serving capacity, benchmarks, and pricing are still being worked out. We’re talking with potential partners to understand where Creda can be useful.

Have a specific model or workload in mind? We’d like to hear about it.

hello@withcreda.com