Fast web search with LLM-powered answers. Returns direct responses for factual queries or detailed summaries with citations for open-ended questions.
Added Dec 23, 2025
Context Window
N/A
Max Output
4.1K
Pricing
Fixed cost: $0.005
Token Pricing
N/A
Benchmarks
Benchmarks
Performance metrics and benchmarks
No benchmark data is available yet for this model.
Providers
Auto routing is available for this model. Explicit provider selection is not available.
Loading provider options…
Related text models
Compare Exa (Answer) with similar models from the same provider or model family.
Web Answer
fastgptFast web answers with citations, tuned for factual questions and quick summaries.
Abliterated Model Large V2
abliteration-ai/abliterated-model-large-v2Abliteration.ai's default unrestricted large text reasoning model is derived from GLM-5.3 and weight-modified to reduce refusals compared with the base model. It is intended for harder reasoning and evaluation workloads, with automatic prompt caching and a one-million-token context window.
Granite 4.2 8B
ibm-granite/granite-4.2-8bIBM Granite 4.2 8B is an Apache 2.0-licensed dense model with native step-by-step reasoning and specialized training for agentic work. It can plan before acting, sequence tools, navigate codebases, work in terminals, and verify results across coding, search, mathematics, science, and complex instruction-following tasks.
GLM 5.3 TEE
TEE/glm-5.3GLM-5.3 is Z.AI's open-weight reasoning model for complex software engineering, autonomous agents, vulnerability research, and long-horizon tasks. This text-only deployment runs inside a Phala Trusted Execution Environment with Redpill attestation and signed completion receipts.
GLM 5.3 Flash TEE
TEE/glm-5.3-flashGLM-5.3 Flash is Z.AI's natively multimodal 320B MoE reasoning model with 18B active parameters, served by Phala inside a Trusted Execution Environment with Redpill attestation and signed completion receipts.
Tencent Hy4 Preview
tencent/hy4-previewHy4 Preview is Tencent's 770B-parameter mixture-of-experts model with 49B active parameters. It is designed for coding agents, complex tool-use workflows, and productivity tasks, with a 1M-token context window and configurable reasoning effort.