DS1 Embedding Model - Code
Takara · Software Development
Certification per AWS Marketplace.
Evidence tier Source Confirmed · 4 captures on record
What the publisher says
As described on AWS Marketplace.
Takara's ds1-code is a CPU-based embedding model purpose-built for understanding code at the semantic level, underpinning Miru, a code-retrieval infrastructure layer for AI coding agents. Rather than dumping entire code bases into an LLM's context window and hoping it identifies what's relevant, Miru combines ds1-code with hybrid retrieval - BM25, lexical search, vector search, and ranking - to deliver the exact function, module, or dependency chain an agent needs in a single retrieval step. Because ds1-code runs on CPU rather than GPU, it avoids GPU procurement, queuing, reservation planning, and regional rollout complexity, while scaling close to developers globally. The result is over 50% token reduction and 40%+ lower token costs across coding agent workflows, with retrieval tasks up to 60% faster and less turns required than conventional approaches, without any impacting the quality of generated code. Miru can be tried via a public API to benchmark value against real code generation workflows, and deployed as a SageMaker endpoint in a customer's own AWS environment - keeping source code and software IP inside a controlled environment.
NOTE: ds1-code is designed to be exclusively used with Miru (https://github.com/takara-ai/miru-code)
Highlights
Highlighted by the publisher on AWS Marketplace.
Cut coding agent token spend. Modern AI coding agents burn most of their token budget not on writing code, but on finding it - reading files, re-loading previously seen context. Miru, powered by ds1-code, replaces that brute-force search with precise semantic retrieval, delivering exactly the code an agent needs in a single step, cutting token consumption across real coding workflows. It works underneath the agents teams already use, including Claude Code, Cursor, Codex, and GitHub Copilot.
CPU-based architecture that scales without GPU bottlenecks. ds1-code is built to run entirely on CPU, eliminating the need for GPU reservation that typically constrain AI infrastructure rollouts. Letting Miru scale without the lead times or cost overhead of GPU-based retrieval, while still delivering search that is up to 60% faster than conventional approaches. This means a lower-cost path to giving every coding agent fast, high-quality retrieval, wherever developers are located.
Sovereign deployment for code and IP that must stay under your control. Source code is one of an organisation's most sensitive assets. Try Miru instantly via the public API to benchmark real workflows, then deploy as a SageMaker endpoint in your AWS environment. Code and IP stay inside your controlled environment at every stage, giving CIOs and security teams a sovereign deployment model that meets data residency needs without giving up the speed or cost advantages of ds1-code.
Agent build and provenance
See the full provenance
The layer-by-layer build, the evidence behind each claim, the risk basis and the cross-marketplace links are open to any account. Some rows are disclosed, some the source leaves Unknown; a free account shows you which.
Plans and pricing as listed
13 listed- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
- HostHrs
Refund terms
As stated by the publisher on AWS Marketplace.
Refunds are furnished in line with the EULA only. Please contact support@takara.ai for assistance.
Sources
Linked repositories
1 repoUnknown means this listing does not publish a repository. It is not a statement that the code is closed, and a linked repository is not a claim that the publisher wrote it: the registry computes that relationship privately and does not publish it.
Evidence risk is the share of the build you cannot see before you deploy, not a security rating. Sign in to see the layer-by-layer basis for this band.

