meta-llama/llama-3.1-8b-instructMeta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient. It has demonstrated strong performance compared to...
Billing follows the provider's own reported cost rather than the catalogue price — cached input is cheaper, and reasoning tokens bill as output. Your usage page is the record of what you actually paid.
One free message per account, no deposit needed. The reply is length-capped — it's a taste, not a quota.
Connect a wallet and sign in to use your free try.
Connect a walletEvery model goes through the same endpoint — just use this page's ID as model.
curl https://macdecloud.com/v1/chat/completions \
-H "Authorization: Bearer dcld-sk-..." \
-H "Content-Type: application/json" \
-d '{"model":"meta-llama/llama-3.1-8b-instruct","messages":[{"role":"user","content":"Hello"}]}'Add "stream": true for token-by-token output over standard SSE.
curl -N https://macdecloud.com/v1/chat/completions \
-H "Authorization: Bearer dcld-sk-..." \
-H "Content-Type: application/json" \
-d '{"model":"meta-llama/llama-3.1-8b-instruct","stream":true,"messages":[{"role":"user","content":"hi"}]}'This model supports tools. Replies may carry tool_calls; run them and append the results as role: "tool" messages, then call again.
{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [{ "role": "user", "content": "What's the weather in Paris?" }],
"tools": [{
"type": "function",
"function": {
"name": "get_weather",
"parameters": {
"type": "object",
"properties": { "city": { "type": "string" } },
"required": ["city"]
}
}
}]
}Supports response_format, so the model returns JSON matching a schema you give it instead of prose you have to parse.
{
"model": "meta-llama/llama-3.1-8b-instruct",
"messages": [{ "role": "user", "content": "..." }],
"response_format": {
"type": "json_schema",
"json_schema": {
"name": "result",
"schema": {
"type": "object",
"properties": { "answer": { "type": "string" } },
"required": ["answer"]
}
}
}
}