Companies

Datumo

datumo.com

Datumo builds Big Data and cloud platforms and evaluates LLMs for business customers.

HQWarsaw, Masovian Voivodeship, Poland
Employees201-1000
Funding$28.7M
2 active roles
Profile 6mo agoJobs checked 21h ago
AI / MLAI ApplicationB2B SaaSSeries B$10M-$50M

About

Datumo is a Poland-based data and AI company that builds Big Data and cloud platforms, provides platform-engineering consulting, and offers Datumo Eval for generating datasets and evaluating LLM-powered services. It sells primarily to business and enterprise customers, differentiating through end-to-end data-platform expertise and experience across more than 250 million data tasks.

Market

Datumo competes in the data-centric AI infrastructure market, spanning AI training data, LLM evaluation, reliability monitoring, and AI safety validation. Its positioning is a full-stack, data-led evaluation platform: unlike tools focused primarily on tracing or scoring existing model interactions, Datumo generates domain-specific query-ground-truth datasets from customer documents, evaluates responses at the claim level, and automates red teaming and risk discovery. Its differentiation is reinforced by proprietary dataset expertise, more than 250 million data tasks, Korean trustworthiness benchmarks, and industry-specific evaluation workflows.

Target Customers

Datumo primarily targets large enterprises, financial institutions, and generative-AI providers deploying customer-facing or high-risk LLM applications, including organizations in e-commerce, finance, education, legal, healthcare, and customer service. Likely buyers are AI/ML, product, engineering, model-risk, and AI-safety teams that need evidence-based reliability and safety validation before deployment.

At a Glance

Problem

Building and operating LLMs and LLM-powered services creates a quality and reliability problem: teams need to know whether a model is answering the right questions, producing useful domain-specific results, and improving as it is tuned. Datumo addresses this evaluation bottleneck, where weak or poorly aligned test sets can make model comparisons unreliable and allow quality problems to reach users. The practical economic pain is wasted engineering effort, evaluation cycles, and model-serving resources, with the highest-value use case being reliable testing of LLMs and AI services before and after deployment.

The core problem is especially acute when evaluation must reflect a particular domain, user intent, or business standard rather than generic benchmark performance. Datumo’s emphasis on question quality, intent alignment, and reliability evaluation suggests that its target customer is an organization that needs repeatable, domain-relevant evidence that its AI systems are working as intended.

Product / Service

Datumo’s main product is Datumo Eval, an agentic evaluation workflow that automatically assesses question quality and returns domain-specific, intent-aligned results. It can automatically generate “golden” question sets using either built-in or custom metrics, then help teams evaluate and enhance their LLMs and LLM-powered services. This turns evaluation from a largely manual test-design exercise into a repeatable software-assisted process.

The company is positioned as a data-centric AI provider specializing in AI training-data services and reliability evaluation. The benefit is faster, more consistent, and more customized testing, while allowing customers to apply their own quality criteria instead of relying only on generic benchmarks. Public descriptions also associate Datumo with data engineering and cloud-computing consulting, suggesting a broader services capability around data and AI infrastructure, although Datumo Eval is the clearest current product expression on datumo.com.

Market

Datumo competes in the emerging market for LLM evaluation, AI quality assurance, reliability testing, and AI training-data infrastructure. Its differentiation, based on the available product description, is the combination of automated golden-set generation, custom evaluation metrics, domain specificity, and agentic assessment rather than a simple generic chatbot or one-off consulting engagement.

The supplied Datumo sources do not identify named competitors or provide quantified customer, revenue, usage, or deployment traction. A company profile describes Datumo as unfunded, so the most supportable characterization is an early-stage or privately operated company with no publicly evidenced funding or commercial scale in the available material—not definitively pre-revenue. Its public positioning is nevertheless clear: it is building an evaluation and data-services layer for organizations developing and deploying LLM-based products.

Founders & Leadership

Daniel PogrebniakFounder
Co-founder and Data Engineer
Piotr GuzikCEO
Michał MisiewiczChief Technology Officer
Rafał RozpondekCDO
Julia KasparekCOO

Funding History

2025-08
Series B$15.5M

Salesforce Ventures

Recent News

2026-05-28product
MinT: A New Way to Manage LLM Versions

Datumo introduced MinT, an infrastructure approach that uses LoRA adapters as the operational unit across training, evaluation, serving, rollout, and rollback. The company says the approach can substantially reduce training-to-serving handoff time compared with moving full model checkpoints.

2026-04-10partnership
[MWC 2026] AI Red Team Challenge

Datumo reported co-hosting the Global AI Red Team Challenge with GSMA for the second consecutive year at MWC 2026. It also showcased the Datumo Platform and Dataset Store, with support from SK Telecom and interest from major international telecom operators.

2026-03-16partnership
Selectstar (Datumo) Enters GSMA Telecom AI Alliance as Korea’s Only Startup

Datumo, also known as Selectstar, joined the GSMA-led Open Telco AI alliance, a group of more than 40 telecommunications, technology, and academic organizations. The company also participated in the alliance’s telecom-focused Open Telco Benchmark and MWC 2026 red-team activities.

2025-12-04product
How to ‘Actually’ Evaluate LLMs

Datumo published a product-focused explanation of Datumo Eval, describing automated evaluation-data generation, claim-level analysis, and AI red teaming for testing model reliability before deployment.

2025-10-21partnership
“Building a More Reliable AI Ecosystem” – Interview with Datumo CEO David Kim

SK Telecom interviewed Datumo CEO David Kim about the company’s role in the K-AI Alliance and SKT’s Sovereign AI Foundation Model consortium. Datumo said it was leading the data domain and contributing dataset creation and reliability verification through Datumo Eval.

2025-08-11funding
Salesforce Ventures Joins Korean AI Startup Datumo’s Series B

Forbes reported that Datumo closed a $15.5 million Series B led by investors including Salesforce Ventures, ACVC Partners, SBI Investment, KB Investment, Shinhan Venture Investment, Kiwoom Investment, and Moorim Capital. The investment was described as Salesforce Ventures’ first investment in a South Korean AI startup and could support collaboration around AI agents and customers.

2025-08-11funding
Seoul-based Datumo raises $15.5M to take on Scale AI, backed by Salesforce

TechCrunch covered Datumo’s $15.5 million financing, which brought its reported total raised to approximately $28 million. The company said it would use the funds to accelerate automated enterprise-AI evaluation tools and expand go-to-market operations in South Korea, Japan, and the United States.

2025-08-04partnership
SKT Consortium Selected for MSIT Project, Aiming to Pioneer Korea’s Leading Proprietary AI Foundation Model

SK Telecom announced that its consortium, including Selectstar (Datumo), was selected as a core team for South Korea’s Ministry of Science and ICT Proprietary AI Foundation Model project. Datumo’s role covers data reliability technology, dataset creation through CashMission, and model stability evaluation using Datumo Eval.

Active Roles

2
Fully Remote/Data & Analytics/195d ago
Fully Remote/Data & Analytics/197d ago

Business Model

Datumo generates B2B revenue from data-engineering, Big Data, cloud-platform, migration, and AI implementation consulting, alongside its Datumo Eval software platform for LLM dataset generation and evaluation. Public sources reviewed do not disclose specific pricing or packaging.

Products

Datumo Eval automated LLM reliability and trustworthiness evaluation platformIndustry-specific evaluation datasets and golden question setsAI training-data services and datasets, including premium and open datasetsAutomated AI red-teaming and safety evaluation capabilities

Customers

SamsungSK Group

Tech Stack

Large language models (LLMs)Multi-agent/agentic workflowsDocument-grounded synthetic evaluation-data generationCustom-metric and claim-level evaluationAutomated AI red teaming

Competitors

Braintrust
Arize
Maxim AI
LangSmith
Langfuse
MLflow

Key Investors

Salesforce