Agent

Model Configurations

Pick the right brain and voice for the job, then tune latency, cost, and creativity.


LLM models

ModelBest forLatencyCost
Claude 3.5 SonnetComplex reasoning, long contextMediumHigher
GPT-4oGeneral purpose, fast responsesLowMedium
GPT-4o MiniSimple agents, high volumeVery lowLow

Voice models

ProviderVoicesBest for
ElevenLabs50+Most natural, wide emotional range
PlayHT30+Fast and cost-effective at volume

Temperature

  • 0.0 โ€” Deterministic โ€” same input, same output. Good for structured calls.
  • 0.7 โ€” Creative and varied. Good for conversational agents.
  • 1.0 โ€” Very creative. Use sparingly.

Max tokens

Default is 150. Increase it for agents that give long explanations, decrease it for simple booking agents that should stay brief.