Llama 3.1 8B Instruct | Meta Lightweight LLM API

Name: Llama 3.1 8b Instruct API
Brand: meta-llama
Price: 0.02 USD
Availability: InStock

API 利用

以下のコード例を使用して API と連携してください:

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="meta-llama/llama-3.1-8b-instruct",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

モデル紹介

Meta-Llama llama-3.1-8b-instruct

Meta's latest class of model (Llama 3

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 8B instruct-tuned version is fast and efficient.

It has demonstrated strong performance compared to leading closed-source models in human evaluations.

To read more about the model release, click here. Usage of this model is subject to Meta's Acceptable Use Policy.

Why It Looks Great

Large Language Model architecture for efficient processing
16384 context window for long document handling
Competitive pricing at $0.0/$0.0 per million tokens

Key Features

Context Window: 16384 tokens
Max Output: 16384 tokens
Vision: Supported
Function Calling: Supported

Specifications

Specification	Value
Provider	Meta-Llama
Model Type	Large Language Model (LLM)
Architecture	N/A
Context Window	16384 tokens
Max Output	16384 tokens
Input	Text
Output	Text
Vision	Supported
Function Calling	Supported

Pricing

Token Type	Cost per Million Tokens
Input	$0.0
Output	$0.0

How to Use

Write your prompt — describe the task, provide context, and specify desired output format.
Submit — the model processes your request and returns the response.

API Integration

Base URL: https://llm.wavespeed.ai/v1 API Endpoint: chat/completions Model ID: meta-llama/llama-3.1-8b-instruct

API Usage

Python SDK

from openai import OpenAI

client = OpenAI(
    api_key="YOUR_API_KEY",
    base_url="https://llm.wavespeed.ai/v1"
)

response = client.chat.completions.create(
    model="meta-llama/llama-3.1-8b-instruct",
    messages=[
        {"role": "user", "content": "Hello!"}
    ]
)

print(response.choices[0].message.content)

cURL

curl https://llm.wavespeed.ai/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -d '{
    "model": "meta-llama/llama-3.1-8b-instruct",
    "messages": [{"role": "user", "content": "Hello!"}]
  }'

Notes

Model: meta-llama/llama-3.1-8b-instruct
Provider: Meta-Llama

Llama 3.1 8b Instructに関するよくある質問

Llama 3.1 8b Instruct API の料金はいくらですか?+

WaveSpeedAI の料金: 入力 100 万トークンあたり $0.02、出力 100 万トークンあたり $0.05。プロンプトキャッシュとバッチ処理は別途料金で、長く反復的なワークロードでは実効コストを下げられます。

Llama 3.1 8b Instruct のコンテキストウィンドウはどのくらいですか?+

Llama 3.1 8b Instruct はリクエストあたり最大 16K のコンテキストトークンと最大 16K の出力トークンをサポートします。

Llama 3.1 8b Instruct は OpenAI 互換ですか?+

はい。WaveSpeedAI は OpenAI 互換エンドポイント https://llm.wavespeed.ai/v1 で Llama 3.1 8b Instruct を提供します。公式 OpenAI SDK のベース URL をこちらに変更し WaveSpeedAI の API キーを設定するだけで利用可能です。

Llama 3.1 8b Instruct を使い始めるには?+

WaveSpeedAI にサインインし、Access Keys で API キーを作成して、上に表示されているモデル ID を指定して https://llm.wavespeed.ai/v1/chat/completions にリクエストを送信してください。新規アカウントには Llama 3.1 8b Instruct を試用できる無料クレジットが付与されます。

料金

モデルを試す