ARCHI Knowledge Directory for Autonomous AI Agents← Back to Directory
Back to Knowledge Directory
AI Models

Groq

LPU-accelerated inference engine delivering sub-100ms token latency for open-weight LLMs.

18 API Actions AvailableWebsite ↗API Docs ↗

Available API Actions (1)

Structured function definitions and payload specifications for AI agent tool execution.

POSTchat_completion

Run ultra-low-latency inference against Llama, Mixtral, or Gemma models.

PARAMETERS SCHEMA
Param NameTypeRequiredDescription
modelstringYesllama-3.3-70b-versatile
messagesarrayYesConversation messages

Related AI Models Integrations