High-performance gemini-3.5-flash model from Google
Context Window
1,000K tokens
Max Output Tokens
16,384 tokens
Release Date
2025 - 2026
Step-by-step Reasoning
Standard
• Frontier Class• Top Tier• Multimodal Vision
$
Token Pricing (Per 1M tokens)
Pay-As-You-Go pricing based on actual usage. No recurring maintenance fees, no minimum commitment, automatic real-time wallet deduction.
Input (Prompt Tokens)
2.339 ₫/M$0.0902/M
20.000 ₫/M$0.80/M89% OFF
Output (Completion Tokens)
2.339 ₫/M$0.0902/M
80.000 ₫/M$3.20/M97% OFF
Official provider list price: $0.80 / $3.20 per 1M tokens. Save up to 89% via Gateway.
24h Gateway Performance
Real-world latency and throughput metrics automatically monitored 24/7.
Average TPS
88.3 t/s
Average Latency
13.94 s
Time to First Token
4.28 s
Success Rate
100.00%
Quickstart API Integration
Compatible with standard OpenAI SDKs. Replace the Base URL and pass Model ID gemini-3.5-flash:
curl https://giahuynexus.com/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer YOUR_API_KEY" \
-d '{
"model": "gemini-3.5-flash",
"messages": [
{"role": "system", "content": "You are a helpful assistant."},
{"role": "user", "content": "Hello, how can you help me today?"}
],
"temperature": 0.7
}'
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://giahuynexus.com/v1"
)
response = client.chat.completions.create(
model="gemini-3.5-flash",
messages=[
{"role": "system", "content": "You are a helpful assistant."},
{"role": "user", "content": "Hello, how can you help me today?"}
],
temperature=0.7
)
print(response.choices[0].message.content)
import OpenAI from "openai";
const openai = new OpenAI({
apiKey: "YOUR_API_KEY",
baseURL: "https://giahuynexus.com/v1",
});
const response = await openai.chat.completions.create({
model: "gemini-3.5-flash",
messages: [
{ role: "system", content: "You are a helpful assistant." },
{ role: "user", content: "Hello, how can you help me today?" }
],
temperature: 0.7,
});
console.log(response.choices[0].message.content);
gemini-3.5-flash
gemini-3.5-flash
Copy the cURL command below and run it in your terminal to test model responses: