AI Model APIs
AI Model APIsProviders
AI Model APIs

The complete platform for comparing AI models. Find pricing, capabilities, and the perfect model for your use case.

contact@aimodelapis.com

Resources

  • AI Models APIs
  • Providers

About

  • About
  • Contact
  • Privacy Policy
  • Terms of Service

© 2026 AI Model APIs. All rights reserved.

Back to Models
Neon

Llama 3.1 8B Instruct API

Neon
About

Process massive datasets with Llama 3.1 8B Instruct, featuring an expansive 131K context window for long-document analysis. This model delivers cost-effective pricing at $0.15/1M input and $0.45/1M output tokens, native tool calling support, open weights architecture. Access Llama 3.1 8B Instruct via the Neon API with up to 16K output tokens.

Capabilities

Input Modalities

text

Output Modalities

text
Reasoning
No
Structured Output
Yes
Tool Use
Yes
WeightsOpen
Temperature
Adjustable
Attachment
Not Supported
Limits
Context Window
131K

Tokens

Input Limit
131K

Tokens

Max Output
16K

Tokens

Pricing

Standard (per 1M tokens)

Input
$0.15
Output
$0.45
Frequently Asked Questions

Llama 3.1 8B Instruct by Neon costs $0.15 per 1M input tokens and $0.45 per 1M output tokens.

Model Details
ID
meta-llama-3-1-8b-instruct
Provider
Neon
Family
llama
Release Date
Jul 23, 2024
Knowledge Cutoff
Dec 31, 2023
API Integration
NPM Package
@ai-sdk/openai-compatible
Environment Variables
NEON_AI_GATEWAY_BASE_URL
NEON_AI_GATEWAY_TOKEN
API Base URL
${NEON_AI_GATEWAY_BASE_URL}/v1
Documentation