Jamba-Instruct Now Available in Snowflake Cortex AI
Key point
Snowflake has integrated AI21's Jamba-Instruct, which supports a 256K context window, into Cortex AI.
Details
Snowflake and AI21 have partnered to make Jamba-Instruct available for serverless inference in Snowflake Cortex AI. Customers can now instantly access this model with simple SQL statements within the secure Snowflake AI Data Cloud.
Jamba-Instruct supports a context window of up to 256K tokens, enabling it to summarize or extract information from about 800 pages of documents in a single inference. Notably, this model is the world's first commercial LLM to successfully scale a SSM-Transformer hybrid architecture, delivering outstanding performance and cost efficiency without quality degradation even over long contexts.
The long context window enhances enterprise AI use cases in various ways:
- Extensive Document Analysis: Long documents such as 10-K reports, meeting minutes, and clinical trial interviews can be efficiently summarized and analyzed.
- Simplified RAG Architecture: Instead of building complex multi-step RAG architectures to find information scattered across multiple documents, large chunks can be processed at once, improving performance.
- Many-shot Prompting: Numerous examples can be included in the prompt to precisely guide the desired style or format.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.