Model Library/Llama3 70B Instruct

Llama3 70B Instruct

meta-llama/llama-3-70b-instruct

Meta's latest class of model (Llama 3) launched with a variety of sizes & flavors. This 70B instruct-tuned version was optimized for high quality dialogue usecases. It has demonstrated strong performance compared to leading closed-source models in human evaluations.

Fonctionnalités

API sans serveur

Documentation

meta-llama/llama-3-70b-instruct is available via Novita's serverless API, where you pay per token. There are several ways to call the API, including OpenAI-compatible endpoints with exceptional reasoning performance.

Déploiements à la demande

Documentation

On-demand deployments allow you to use meta-llama/llama-3-70b-instruct on dedicated GPUs with high-performance serving stack with high reliability and no rate limits.

Sans serveur disponible

Exécutez des requêtes immédiatement, ne payez que pour l’utilisation

Entrée$0.51 / M Tokens

Sortie$0.74 / M Tokens

Utilisez les exemples de code suivants pour intégrer notre API :

1from openai import OpenAI
2
3client = OpenAI(
4    api_key="<Your API Key>",
5    base_url="https://api.novita.ai/openai"
6)
7
8response = client.chat.completions.create(
9    model="meta-llama/llama-3-70b-instruct",
10    messages=[
11        {"role": "system", "content": "You are a helpful assistant."},
12        {"role": "user", "content": "Hello, how are you?"}
13    ],
14    max_tokens=8000,
15    temperature=0.7
16)
17
18print(response.choices[0].message.content)

Infos

Fournisseur

Llama

Quantification

fp8

Fonctionnalités prises en charge

Longueur du contexte

8192

Sortie maximale

8000

Serverless

Pris en charge

Structured Output

Pris en charge

Capacités d’entrée

text

Capacités de sortie

text