Skyfall 31B v4.2 — Skyfall | Parasail
Specs & substance
Spec sheet
| Attribute | Value |
|---|---|
| Developed by | Skyfall |
| Model family | Skyfall |
| Category | compact |
| Modality | Text |
| Context window | 128K tokens |
| Architecture | Dense |
| Version | 4.2 |
| License | Apache 2.0 |
| Pricing | $0.55 in $0.80 out · $0.25 cache |
| Released | Feb 2026 |
| Endpoint | parasail-skyfall-31b-v42 |
Fine-tuned model optimized for creative writing and conversational tasks.
Skyfall 31B v4.2 is part of the Skyfall family by Skyfall, categorized as a compact model. It supports a 128K tokens context window and is built on a Dense architecture.
Key strengths: creative, conversational. Parasail serves Skyfall 31B v4.2 on a global fleet of current-gen GPUs behind a single OpenAI-compatible endpoint — with per-token pricing, no minimums, and dedicated capacity options when you need guaranteed throughput.
Drop-in via the OpenAI SDK
Point any OpenAI-compatible chat client at Parasail and change the model name. That's it.
Python
from openai import OpenAI
client = OpenAI(
api_key="YOUR_API_KEY",
base_url="https://api.parasail.io/v1"
)
response = client.chat.completions.create(
model="parasail-skyfall-31b-v42",
messages=[
{"role": "user", "content": "Hello, what can you do?"}
],
stream=True,
stream_options={"include_usage": True},
top_p=1,
max_tokens=1000,
temperature=1
)
for chunk in response:
if chunk.choices and chunk.choices[0].delta.content is not None:
print(chunk.choices[0].delta.content, end="", flush=True)
Bash
curl https://api.parasail.io/v1/chat/completions \
-H "Authorization: Bearer $PARASAIL_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "parasail-skyfall-31b-v42",
"messages": [{"role": "user", "content": "Hello, what can you do?"}],
"stream": true,
"max_tokens": 1000
}'
Explore the library
Model DeepSeek V4 Flash
Ultra-fast, ultra-cheap reasoning model for high-throughput workloads.
$0.14/M in
Model Qwen3.6 35B-A3B
Efficient MoE model with 3B active parameters — great price-to-performance.
$0.15/M in
Model Qwen3.5 35B-A3B
Compact MoE model with strong reasoning at minimal cost per token.
$0.15/M in
Run Skyfall 31B v4.2 in the platform
Start with free credits — no credit card. Call any model in minutes.