Kimi Deep Research: How It Works, Access Test and Limits

Kimi Deep Research is a research agent that plans a task, searches the web, evaluates sources and produces a cited long-form report. It is designed for work that needs more investigation than an ordinary chat answer, such as literature reviews, market research, policy analysis and multi-source fact finding.

This guide separates three things that are often blended together:

  1. what Kimi officially says the product does;
  2. how to use it without losing control of the research scope; and
  3. what our authenticated access attempt observed and how the future performance test will be run.

Test status: On August 4, 2026, we logged in, selected Deep Research with K3 High, entered a fixed five-question official-source prompt and attempted to start it. Kimi displayed: “Too many people are chatting with Kimi right now. Subscribe to enter a dedicated priority queue!” Research did not begin. We therefore report an access-blocked attempt, not an accuracy, speed, source-quality, citation or credit-use result.

Kimi AI Guide is independent and is not affiliated with Moonshot AI. Vendor-published averages below are labeled as such; they are not our measurements.

What Kimi Deep Research is

Kimi’s current Help Center describes Deep Research as an Agent product powered by Kimi-Researcher. It says the system can clarify the question, plan searches, inspect pages, use tools, reason iteratively and assemble a cited report. The official technical report describes a research agent with parallel search, a text browser and code execution.

The product is distinct from ordinary Kimi Search:

ProductBest fitTypical interactionWhat this page evaluates
Kimi Chat with web searchA current fact or a bounded questionShort conversational answerNot the main subject of this page
Kimi Deep ResearchA question requiring planning, many sources and synthesisAsynchronous research task and long reportEnd-to-end report coverage, factual support and source handling
Kimi AgentCreating or modifying a deliverable with multiple toolsAutonomous task executionCovered separately in the Kimi Agent guide

For a focused benchmark of short search answers, see our Kimi Search Citation Accuracy Test. That test uses many atomic questions; this page examines complete research reports.

Current documented capabilities and limits

The table below records official product statements checked on August 4, 2026. “Vendor-reported” means Kimi published the figure; we have not yet reproduced it.

ItemCurrent official descriptionEvidence status
Underlying product modelKimi-ResearcherOfficially documented
Research workflowClarification, reasoning, active search, iterative review, tool use and report generationOfficially documented
Typical execution time10-25 minutes, asynchronousVendor-reported range; live measurement pending
Context length128K tokensOfficially documented; usable task capacity not independently measured
Typical activityAbout 23 reasoning steps, 74 planned keywords and 206 discovered URLs, with 3.2% retainedVendor-reported averages, not fixed quotas
Text reportMore than 10,000 words on average, with about 26 traceable sourcesVendor-reported averages, not a promise for every task
Visual reportInteractive, shareable HTML report with structured layout and mind mapsOfficially documented; export test pending
Saving resultsText reports can be copied or downloaded as PDF or Word; visual reports can be shared and saved from previewOfficially documented; export test pending
Execution behaviorThe task continues asynchronously; Kimi advises refreshing rather than stopping an apparently idle runOfficially documented
CreditsMembership features share a credit pool and consumption depends on tokens; Kimi gives roughly 5-10% of free-tier credits as an example for a deep research reportVendor estimate; actual consumption varies and will be recorded per run
Simple questionsKimi recommends standard chat for simple Q&AOfficially documented

These figures describe the consumer product, not a Deep Research API. Kimi’s current API model-selection help says a Deep Research API is not available. Do not reuse the 128K Deep Research context figure as an API model limit; Kimi publishes separate API limits for models such as K3 and K2.6.

How to use Kimi Deep Research

1. Choose a question that needs research

Use Deep Research when the answer depends on several sources, disputed evidence, a timeline or a structured comparison. Use ordinary chat when one authoritative page can answer the question directly.

Weak request:

Tell me about the AI market.

Researchable request:

Using sources published from January 2024 through July 2026, compare the
adoption of generative AI by small businesses in the United States and the
United Kingdom. Separate survey results from estimates, identify the sample
and sponsor of every survey, and do not combine metrics with different
definitions. Prioritize government statistics and original survey reports.

2. Define the evidence rules before the task starts

A strong request states:

  • the exact question and intended reader;
  • the geographic and time boundaries;
  • the fields that must be covered;
  • preferred and excluded source types;
  • how conflicting evidence should be handled;
  • what must be marked not found rather than inferred; and
  • the required format and length.

Use this reusable prompt frame:

Research question: [one bounded question]
Audience: [who will use the report]
Cutoff: Use information published by [date and time zone].
Scope: [countries, products, years or documents]
Required fields: [list every field]
Source policy: Prioritize [primary/official/peer-reviewed sources].
Do not use: [aggregators, anonymous posts, search snippets, etc.]
Conflict rule: Show conflicting figures side by side; do not average them.
Missing-data rule: Write "not found" and describe the search performed.
Citations: Cite every material factual claim at the point where it appears.
Output: [table/report/appendix], no more than [length].

3. Use the clarification stage

Kimi says Deep Research asks follow-up questions before searching. Treat this as a scope checkpoint. Correct an unwanted geography, date range, definition or deliverable here. A long clarification answer can introduce new ambiguity, so answer with short, explicit constraints.

4. Let the asynchronous task finish

Kimi documents a typical 10-25 minute run and says the task can continue while you leave the conversation. If the page appears unchanged, refresh it rather than clicking Stop output. The official FAQ says manually stopping or closing a running task can still consume credits.

5. Audit the report before using it

Do not judge a report by its length, citation count or design. Check:

  1. whether every requested field was answered;
  2. whether each citation opens;
  3. whether the cited passage supports the nearby claim;
  4. whether a primary source was available but ignored;
  5. whether dates and units match the source;
  6. whether estimates are labeled as estimates;
  7. whether contradictory sources were disclosed; and
  8. whether recommendations go beyond the evidence.

Observed access attempt: August 4, 2026

We performed a signed-in access attempt before assigning any performance claim.

FieldObserved state
Account stateLogged in; subscription tier is not inferred from the blocker
Product modeDeep Research selected
Displayed modelK3 High
InputFive-question official-sources-only task scope; the complete verbatim input was not independently retained
Start resultBlocked before research execution
Visible message“Too many people are chatting with Kimi right now. Subscribe to enter a dedicated priority queue!”
Search activityNone observed; the research workflow did not start
Report or citationsNone generated
Performance scoringNot applicable

The queue message demonstrates the access state for this account at that moment. It does not show that Deep Research failed the questions, that paid access always works, that free access never works, or that the documented 10-25 minute execution range is inaccurate.

Normalized record of the submitted task scope

The priority-queue modal obscured part of the input in the retained screenshot. The block below is a normalized reconstruction of the verified five-question scope and constraints used for the attempt; it is not presented as a complete screenshot-visible verbatim transcript.

Create a concise research report answering these five questions. Use only
official Kimi Help Center, platform.kimi.ai documentation, Moonshot AI
technical pages, or official model cards - never search snippets or
third-party sources. Answer each item in a table with the exact official
value, a short caveat, and the direct source URL per row:

1. What context limit is documented for Kimi Deep Research?
2. What are the documented per-file size and files-per-session limits in
   Kimi Chat?
3. What context windows are documented for the kimi-k3, kimi-k2.7-code and
   kimi-k2.6 API models?
4. Are Project files fully injected into every turn?
5. Do Kimi Membership, Kimi Code and Kimi API share one balance?

Write "not published" when an official source does not publish a requested
value. Keep the report under 300 words and add a limitations section. Record
the access date as August 4, 2026.

The send attempt produced the priority-queue message before a research plan, search query, visited URL or answer appeared. No answer row can therefore be scored.

Kimi Deep Research with K3 High showing a priority-queue subscription message before the five-question research task started.
Authenticated Kimi Deep Research access attempt on August 4, 2026. The priority-queue message appeared before research began; this is access evidence, not a performance result.

Planned independent performance test

The access attempt above is complete. The performance methodology below is still planned and will be executed only when the research task can start. A retry receives a new run ID; it does not overwrite the blocked attempt.

We plan two complementary tasks. DR-01 uses the same five-question official-source audit scope with a known answer key. DR-02 is an external public-data synthesis task. Together they can reveal different failure modes without pretending that two reports establish universal performance.

Fixed environment rules for a future run

For both tasks:

  • start a new English-language Deep Research conversation;
  • use the same authenticated account and record the plan without revealing identity;
  • record the exact product and model labels visible in the interface;
  • do not upload files or provide source URLs in the prompt;
  • submit the frozen prompt without adding hints from the answer key;
  • do not edit, regenerate or steer the report after execution begins;
  • retain every failure and off-topic source, not only the best output;
  • capture the start and completion timestamps in UTC;
  • record credits before and after using the smallest visible unit; and
  • stop and classify the run separately if login, queue access or sufficient credits are unavailable.

DR-01: Five-question Kimi documentation audit

Question: Can Deep Research return the five requested product facts from the correct official consumer or API scope, preserve caveats and cite a direct source for every row?

The future run will reuse the submitted prompt printed above.

DR-01 ground-truth ledger

This ledger was prepared independently of any model output. Every source and value must be rechecked immediately before a future scored run because Kimi documentation can change.

QuestionGround truth at protocol freezePrimary official source
Deep Research context128K tokensDeep Research FAQ
Kimi Chat file limitsUp to 100 MB per file and up to 50 files per sessionKimi overview
API context windowskimi-k3: 1M tokens; kimi-k2.7-code: 256K tokens; kimi-k2.6: 256K tokensKimi Platform introduction and model list
Project file injectionNo. Kimi documents Project files as read on demand rather than fully injected into every turnProjects
Product balancesThe three products do not share one universal balance. Current membership rules include Kimi Code in the membership credit pool, while Open Platform API balance and keys remain separateMembership credit rules and API troubleshooting

DR-02: External primary-source synthesis

Question: Can the product extract closely related estimates without changing years, units, confidence language or population groups?

Exact prompt:

Using only World Health Organization primary publications available by
23:59 UTC on August 4, 2026, write an evidence brief on the global malaria
burden in 2024 as reported in the World malaria report 2025.

Report: estimated global cases and deaths for 2024; the comparable 2023
figures; change in cases; the African Region's share of cases and deaths;
the share of African-region malaria deaths among children under five; the
number of countries using malaria vaccines in routine programmes; the number
of countries and children reached by seasonal malaria chemoprevention; and
WHO's estimate of cases and deaths averted by wider use of new tools in 2024.

Preserve units, years and words such as "estimated." Distinguish report
estimates from reported surveillance counts. Use the report, executive
summary, annexes or WHO release rather than third-party summaries. Cite every
number at the point of use and include page, table or section when available.
If two WHO pages disagree, show both values and do not silently select one.
Keep the brief between 1,000 and 1,500 words and end with a claim-to-source
appendix.

Fixed clarification response:

Keep the scope global and limited to the requested 2023-2024 comparison.
Do not provide medical advice, country rankings or projections. Prioritize the
World malaria report 2025 and its annexes over later summaries, while listing
any material discrepancy you find.

DR-02 ground-truth ledger

FieldGround truth at protocol freezePrimary official source
2024 global cases282 million, estimatedWHO 2025 report release
2023 global cases273 million, implied by WHO’s stated increase of about 9 million; verify against report table before scoringWHO 2025 executive summary
2024 global deaths610,000, estimatedWHO malaria fact sheet
2023 global deaths598,000, estimatedWHO malaria fact sheet
African Region share94% of cases and 95% of deaths in 2024WHO 2025 executive summary
Children under five75% of deaths in the African RegionWHO 2025 executive summary
Vaccine programmes24 countries had introduced malaria vaccines into routine immunization programmesWHO 2025 report release
Seasonal chemoprevention20 countries; 54 million children reached in 2024WHO 2025 report release
Wider-tool estimate170 million cases and 1 million deaths averted in 2024WHO 2025 report release

The 2023 cases value must be checked in the full report immediately before scoring. A value inferred by subtraction is not treated as a report quotation.

Scoring rules

We report separate metrics rather than hiding them inside one star rating.

Required-field coverage

Each requested field receives:

  • 2 – present, unambiguous and in the required format;
  • 1 – present but incomplete, ambiguous or outside the requested format;
  • 0 – omitted or replaced by a different fact.

Coverage = points earned / maximum possible points x 100

Claim-level citation audit

Before opening any cited source, split the report into atomic, externally checkable claims. A sentence containing two dates and one quantity can create three claim rows.

Each claim is then scored:

Dimension210
Citation completenessCitation clearly attached to the claimCitation placement is ambiguousNo citation
EntailmentSource supports the whole claimSource supports only partSource does not support it or cannot be located
Source qualityRequested primary sourceCredible secondary source where primary was requestedUnsuitable, anonymous or search snippet
FreshnessCurrent for the stated cutoff and reference periodDate unclear or newer source was availableStale for the claim
Numeric fidelityValue, unit, population and period all matchOne non-material qualifier is missingMaterial mismatch

We also report these counts separately:

  • fabricated or malformed URLs;
  • links that do not open for a normal reader;
  • unsupported quotations;
  • undisclosed source conflicts;
  • estimates presented as observed counts; and
  • claims added outside the requested scope.

No overall product grade

Two reports cannot establish general accuracy or reliability. We publish the coverage and citation measures separately, show the raw rows and describe failures. We do not convert these results into stars, a universal percentage or a claim that Kimi Deep Research is “best.”

Observed result and planned runs

RunTaskDate UTCResearch startedDurationCredits usedSources citedCoverageCitation entailmentStatus
DR-ACCESS-01Five-question official-source prompt access attemptAugust 4, 2026NoNot applicableNot measuredNot applicableNot scoredNot scoredAccess blocked by priority queue before execution
DR-01-R1Future scored five-question auditPendingPendingPendingPendingPendingPendingPendingPlanned
DR-02-R1Future WHO evidence synthesisPendingPendingPendingPendingPendingPendingPendingPlanned

Not scored is not zero. The blocked attempt produced no research output to evaluate. A future completed run will be added as a new row rather than replacing DR-ACCESS-01.

Raw-data tables

Run metadata

Run IDProtocol versionAccount stateInterface labelModel label shownLanguageBrowser/OSTest regionStart creditsEnd creditsExport obtainedEvidence bundle
DR-ACCESS-011.0 access attemptAuthenticated; tier not assertedDeep ResearchK3 HighEnglishBrowser session; version not recordedNot independently determinedNot recordedNot recordedNoOriginal priority-queue screenshot shown above
DR-01-R11.0PendingDeep ResearchPendingEnglishPendingNot independently determinedPendingPendingPendingPlanned
DR-02-R11.0PendingDeep ResearchPendingEnglishPendingNot independently determinedPendingPendingPendingPlanned

Required-field results

RunField IDRequired fieldOutput valueGround-truth valueCoverage 0-2CorrectCitation IDsNotes
DR-ACCESS-01AllFive requested answersNo output generatedSee frozen ledgerNot applicableNot scoredNonePriority-queue message appeared before research began

Claim-to-source audit

RunClaim IDExact atomic claimCitation URLLink opensSource dateCompleteness 0-2Entailment 0-2Quality 0-2Freshness 0-2Numeric fidelity 0-2/NADecision note
DR-ACCESS-01NoneNo output claim existsNoneNot applicableNot applicableNot applicableNot applicableNot applicableNot applicableNot applicableAccess-only event

Execution timeline

RunUTC timeStageVisible query, URL or eventObservationScreenshot ID
DR-ACCESS-01August 4, 2026; exact UTC time not recordedAccess attemptDeep Research selected; K3 High displayed; send attemptedPriority-queue subscription message appeared before executionkimi-deep-research-priority-queue-2026-08-04.png

The completed tables should be downloadable as CSV in the Kimi AI Test Lab alongside a checksum or immutable file version.

How to verify a Kimi Deep Research report yourself

Use a small claim ledger rather than reading citations casually:

ClaimCitation opens?Passage supports it?Source is primary?Date fits?Your decision
Copy one factual claimYes/NoFull/Partial/NoYes/NoYes/NoKeep/Correct/Remove

Open the destination page, search for the exact number or phrase, inspect the surrounding qualifier and confirm that the page is about the same period and population. A real link can still be a bad citation.

For decisions involving health, law, finance or safety, use the report as a research aid and review the primary material with an appropriately qualified professional.

Practical limitations

The observed result is access-only

The August 4 attempt ended at a queue gate before the research workflow began. It cannot answer whether Kimi would find the correct values, use the requested source scope, finish within the documented time range or consume a particular amount of credit. It also cannot establish general availability: queue conditions and subscription entitlements can change.

One run is a case study, not a reliability rate

Deep Research outputs and live web results can vary. The two planned performance runs can expose concrete successes and failures under dated conditions, but they cannot estimate how often the product will succeed across every subject.

The web changes after the test

A linked page can be revised, moved or removed. We record access dates and retain a permissible excerpt, screenshot or checksum so a later reader can understand what was scored.

Source count is not source quality

Twenty citations can repeat the same press release or fail to support the nearby text. We score claims and source passages, not just the total number of links.

Product labels and plans can vary

Availability, model labels, credits and exports may depend on account, rollout, region and plan. We report only the tested account state and do not generalize it to every user.

The interface does not expose every internal action

Visible queries and URLs are observations from the product interface. They may not constitute a complete internal trace. We describe them as visible activity, not a full chain of reasoning.

Ground truth can also be wrong or ambiguous

Official sources can conflict. When they do, the evaluator records each value and applies the predeclared source hierarchy instead of silently picking the convenient answer.

Frequently asked questions

Did Kimi Deep Research complete our test?

No. We reached the signed-in Deep Research interface, selected K3 High and attempted to submit the fixed prompt, but a priority-queue subscription message appeared before research began. We classify this as access blocked and publish no performance score.

Is Kimi Deep Research free?

Kimi says Deep Research is available with a limited free allowance and that membership features use a shared credit pool. The amount used depends on the task and token consumption; Kimi’s 5-10% figure for a report is explicitly a rough free-tier example, not a fixed price.

How long does Kimi Deep Research take?

The official typical range is 10-25 minutes. That is vendor documentation, not our current measured result. Actual time can depend on the task, service state and account.

How many sources does it use?

Kimi reports about 26 traceable sources on average. An average is neither a minimum nor evidence that every citation is correct. A future completed test will audit each claim-to-source relationship.

Can I leave the page while it runs?

Kimi says the task runs asynchronously and you can return later. Its Help Center recommends refreshing an apparently idle page rather than stopping the task.

Can I use Deep Research through the Kimi API?

Kimi’s current API help says a Deep Research API is not available. The consumer Deep Research product and ordinary Kimi API model calls should be treated as separate products.

Is a cited Kimi report safe to trust without checking?

No. A citation can be stale, partial or unrelated even when it opens. Check important claims against the source, especially for high-stakes decisions.

Related pages


Official sources

Official sources checked August 4, 2026:

  1. Kimi Help Center: Deep Research overview
  2. Kimi Help Center: Deep Research FAQ
  3. Kimi Help Center: Deep Research use cases and prompt library
  4. Kimi Help Center: Membership credit rules
  5. Kimi Help Center: Kimi overview and Chat file limits
  6. Kimi Platform: introduction and model list
  7. Kimi Help Center: Projects
  8. Kimi Help Center: API troubleshooting and product separation
  9. Kimi Help Center: API model selection
  10. Moonshot AI: Kimi-Researcher technical report
  11. WHO: World malaria report 2025 news release
  12. WHO: World malaria report 2025 executive summary
  13. WHO: World malaria report 2025 annexes
  14. WHO: Malaria fact sheet