Model Library/Llama3 70B Instruct
meta-llama/llama-3-70b-instruct

Llama3 70B Instruct

meta-llama/llama-3-70b-instruct
Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 70B instruct-tuned version was optimized for high quality dialogue usecases. It has demonstrated strong performance compared to leading closed-source models in human evaluations.

Fonctionnalités

API sans serveur

Documentation

meta-llama/llama-3-70b-instruct is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Déploiements à la demande

Documentation

On-demand deployments allow you to use meta-llama/llama-3-70b-instruct on dedicated GPUs with high-performance serving stack with high reliability and no rate limits.

Sans serveur disponible

Exécutez des requêtes immédiatement, ne payez que pour l’utilisation

Entrée$0.51 / M Tokens
Sortie$0.74 / M Tokens

Utilisez les exemples de code suivants pour intégrer notre API :

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="meta-llama/llama-3-70b-instruct",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=8000,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Infos

Fournisseur
Llama
Quantification
fp8

Fonctionnalités prises en charge

Longueur du contexte
8192
Sortie maximale
8000
Serverless
Pris en charge
Structured Output
Pris en charge
Capacités d’entrée
text
Capacités de sortie
text