[ Web Proxy ]
URL:
Viewing: https://www.digitalocean.com/data-learning [Back]  [Original]

Data & Learning | DigitalOcean

DigitalOcean Data & Learning

Vector search and RAG infrastructure for your AI agents. Fully managed PostgreSQL, MySQL and Valkey for your production workloads. One platform, no egress fees between them.

Open memory. Fresh data. Systems that get smarter.

Managed Databases

Worry-free database hosting. We offer automated, high-availability setups for MySQL, PostgreSQL, and Redis, removing the need for manual server administration.

Features:

High scalability

Easily scale your database to match business growth. Add CPUs, RAM, and nodes to handle heavier workloads and boost performance. Plus, with automated storage scaling, youll never run out of spaceno manual intervention required.

Fast, reliable performance

Managed Databases run on enterprise-class hardware for fast performance. Run your clusters on Droplets with shared vCPUs or choose Droplets with 100% dedicated vCPUs for mission critical workloads.

End-to-end security

Databases run in your account's private network, and only whitelisted requests via the public internet can reach your database. Data is also encrypted in transit and at rest.

Streamlined setup and maintenance

Launch a database cluster with just a few clicks and then access it via our simplified UI or API. Easily migrate your database from another location with minimal downtime.

Knowledge Bases (GA)

Features:

Diverse sources
Connect to data in S3, Dropbox, and local file system.

Chunking Strategies
Smart defaults out of the box, deep control when you need it. Pick semantic for topical shifts, hierarchical for precise retrieval with broader grounding, section-based for structured docs, or fixed-length for sheer speed - then evaluate, adjust, and re-index until retrieval lands.

Hybrid Search & Advanced Reranking
Enhance retrieval accuracy with sophisticated search techniques that combine keyword and semantic results. Enable bge-reranker-v2-m3 to re-score results with cross-encoder precision. Add a reranking step to your retrieval pipeline. Higher precision on the top results, $0.010 per 1M tokens.

New Embedding Models
Additional open-source models (e5-large-v2, bge-m3) are now available, offering high-precision English retrieval and versatile support for long-form, multilingual documents.

Model Context Protocol (MCP) Support
Turn your Knowledge Bases into a plug-and-play retrieval tool for any MCP-compatible agent framework.

Managed Weaviate (Public Preview)

Skip the complexity of self-hosting and the high cost of vendor lock-in with Managed Weaviate, an open-source compatible solution that lets you focus on building, not operations. Launch production-ready vector infrastructure for your AI apps with 1-click provisioning.

Sign-up for early access

Features:

Simple, predictable pricing
$20/month to start on the Small plan, $120/month Medium, $1,600/month Large - with built in HA functionality, no per-query or per-vector metering.

Built on Weaviate 1.37.1
Full compatibility with the upstream Python, JavaScript/TypeScript, Go, and Java clients, plus REST, gRPC, and GraphQL APIs.

RQ8 compression by default
Rotational Quantization 8-bit is enabled at cluster level, giving roughly 4x less RAM per vector than uncompressed storage while preserving recall.

Native DigitalOcean ecosystem integration
OpenAI and Anthropic-compatible pairing with DigitalOcean Serverless Inference for embeddings.

Integrated Capabilities. One Unified Platform. Scale without Complexity.

Data Thats Ready for your App and your AI

Queryable the moment it lands, on open-source engines PostgreSQL, MySQL, Valkey, Weaviate with fully portable infrastructure. No rewrites to move data in or out.

Native Integration, No Lock-in

Standard APIs and open-source engines across your databases and AI tools. No proprietary SDKs required, no rip-and-replace.

Efficient Unit Economics

No egress fees between your database and your AI stack. Managed databases start at $15/month, Knowledge Bases at $20/month.

Reduced Operational Complexity

We provide a single control panel for your entire AI stack -from raw data storage to operational and vector databases and agent deployment so you dont have to manage multiple vendor relationships, separate billing, and complex authentication layers.

Built for production, Ready for AI

Leverage scalable, low-latency, highly available, manageable and secure managed databases. With our new Advanced Edition, databases can extend the boundaries of scale, failover and manageability.

FAQs about DigitalOcean Data and Learning

Do I need to manage a vector database with Knowledge Bases?
No. Knowledge Bases includes fully managed vector storage and retrieval infrastructure. No need to create a separate vector store.
What are the key capabilities of Managed Weaviate with DigitalOcean?
Our offering includes full Weaviate compatibility (GraphQL, REST, HNSW configuration). DigitalOcean manages backups, patching, and availability. Available now in Public Preview.
Can I use my own models with Knowledge bases?
No. Knowledge bases currently supports 6 embedding models.
How do I access Knowledge Bases?
Directly in the DigitalOcean console under Data Services, via API, via the CLI or through the DigitalOcean AI SDK.
Do Knowledge Bases work with AI agents?
Yes. MCP integration allows any compatible agent framework to connect directly.
How do I access Weaviate capabilities on DigitalOcean?
Weaviate features are in Public Preview.These features can be accessed in theDigitalOcean control panelin the Vector Databases section.
Is reranking required for Knowledge Bases?
No. It is optional but recommended for higher-quality retrieval results.

Web Proxy Viewer  |  New URL  |  Original Page