Elastic Adds Support for Cohere High-Performance Embeddings
Developers can now natively use the Elastic vector database to store and search Cohere’s new int8 text embeddings
April 03, 2024
Share this

Elastic, the company behind Elasticsearch®, announced the Elasticsearch open Inference API now supports Cohere’s text embedding models.

This includes Elasticsearch native support for efficient int8 embeddings, which optimize performance and reduce memory cost for semantic search across the large datasets commonly found in enterprise scenarios.

With this integration, Elasticsearch developers can experience immediate performance gains, including up to 4x memory savings and up to 30% faster search, without impacting search quality.

“We’re excited to collaborate with Elastic to bring state-of-the-art search solutions to enterprises,” said Jaron Waldman, chief product officer at Cohere. “Elasticsearch delivers strong vector retrieval performance on large datasets, and their native support for Cohere’s Embed v3 models with int8 compression helps unlock gains in performance, efficiency, and search quality for enterprise-grade deployments of semantic search and retrieval-augmented generation (RAG)."

“Developers who want to build more intuitive and accurate semantic search experiences for enterprise use cases need to look at Elasticsearch and Cohere,” said Shay Banon, founder & chief technology officer at Elastic. “Innovation is rarely insular, and our work with the great team at Cohere showcases how we bring developers the best of both worlds. The Cohere and Elastic communities now have great models to generate embeddings with support for inference workloads and seamless integration into the leading search and analytics platform that has invested in creating the best vector database.”

Support for Cohere embeddings is available in preview with Elastic 8.13 and will soon be generally available in an upcoming Elasticsearch release.

Share this

The Latest

October 04, 2024

In Part 1 of this two-part series, I defined multi-CDN and explored how and why this approach is used by streaming services, e-commerce platforms, gaming companies and global enterprises for fast and reliable content delivery ... Now, in Part 2 of the series, I'll explore one of the biggest challenges of multi-CDN: observability.

October 03, 2024

CDNs consist of geographically distributed data centers with servers that cache and serve content close to end users to reduce latency and improve load times. Each data center is strategically placed so that digital signals can rapidly travel from one "point of presence" to the next, getting the digital signal to the viewer as fast as possible ... Multi-CDN refers to the strategy of utilizing multiple CDNs to deliver digital content across the internet ...

October 02, 2024

We surveyed IT professionals on their attitudes and practices regarding using Generative AI with databases. We asked how they are layering the technology in with their systems, where it's working the best for them, and what their concerns are ...

October 01, 2024

40% of generative AI (GenAI) solutions will be multimodal (text, image, audio and video) by 2027, up from 1% in 2023, according to Gartner ...

September 30, 2024

Today's digital business landscape evolves rapidly ... Among the areas primed for innovation, the long-standing ticket-based IT support model stands out as particularly outdated. Emerging as a game-changer, the concept of the "ticketless enterprise" promises to shift IT management from a reactive stance to a proactive approach ...

September 27, 2024

In MEAN TIME TO INSIGHT Episode 10, Shamus McGillicuddy, VP of Research, Network Infrastructure and Operations, at EMA discusses Generative AI ...

September 26, 2024

By 2026, 30% of enterprises will automate more than half of their network activities, an increase from under 10% in mid-2023, according to Gartner ...

September 25, 2024

A recent report by Enterprise Management Associates (EMA) reveals that nearly 95% of organizations use a combination of do-it-yourself (DIY) and vendor solutions for network automation, yet only 28% believe they have successfully implemented their automation strategy. Why is this mixed approach so popular if many engineers feel that their overall program is not successful? ...

September 24, 2024

As AI improves and strengthens various product innovations and technology functions, it's also influencing and infiltrating the observability space ... Observability helps translate technical stability into customer satisfaction and business success and AI amplifies this by driving continuous improvement at scale ...

September 23, 2024

Technical debt is a pressing issue for many organizations, stifling innovation and leading to costly inefficiencies ... Despite these challenges, 90% of IT leaders are planning to boost their spending on emerging technologies like AI in 2025 ... As budget season approaches, it's important for IT leaders to address technical debt to ensure that their 2025 budgets are allocated effectively and support successful technology adoption ...