~/models

8 models14 operators onlinepick a model · inspect machines · call via OpenAI SDK

catalog

8 of 8
mlx-community/Qwen2.5-7B-Instruct-4bittext

Advertised on the network by 5 machines. Rates come from each provider's priceList entry for this model id (typically tokens per MTok at the 1:1 uniform rate). Activity is from indexed receipts.

currency · CCfreshest record · just nowlive directory
machines5
input1Mtokens/Mtok
output1Mtokens/Mtok
runs · 7d952req
tokens · 7d2.4M
tokens · 24h682.7K

machines on this model

5 rows

righost24hlast seen

m4-mac-mini · Apple M4, 16 GB

1 req · 1.5K tk

23s ago

Skogs-Mac-mini.local · Apple M4, 16 GB

0 req · 0 tk

just now

Mac Mini 2026 · Apple M4, 32 GB

🛡️ Hardware-attested · experimental

230 req · 652.9K tk

25s ago

μ · Apple M4, 16 GB

10 req · 28.4K tk

just now

M4 · Apple M4, 24 GB

0 req · 0 tk

13s ago

example usage

OpenAI SDK (and curl) — same snippets as API docs.

import OpenAI from "openai";

const client = new OpenAI({
  baseURL: "https://cocore.dev/v1",
  apiKey: "cocore-...",
});

const stream = await client.chat.completions.create({
  model: "mlx-community/Qwen2.5-7B-Instruct-4bit",
  messages: [{ role: "user", content: "Hello" }],
  stream: true,
});

for await (const chunk of stream) {
  process.stdout.write(chunk.choices[0]?.delta?.content ?? "");
}