/generate
Generate
Generate a response using Contextual's Grounded Language Model (GLM), an LLM engineered specifically to prioritize faithfulness to in-context retrievals over parametric knowledge to reduce hallucinations in Retrieval-Augmented Generation and agentic use cases.
The total request cannot exceed 32,000 tokens.
See our blog post and code examples. Email glm-feedback@contextual.ai with any feedback or questions.
post/generate
Request body
Response
Successful Response