Archyvearchyve v2

Introducing archyve v2

A short note on research workflows, paper validation, and turning fragmented scientific literature into actionable intelligence.

August 2026·Product Announcement

archyve is a research-intelligence platform engineered to turn scattered, paywalled, and noisy scientific literature into trustworthy, accessible, and structured research dossiers.

01 / Background

Why archyve exists

The academic web is deeply fragmented. When researchers discover an intriguing paper on IEEE Xplore, Springer Nature, Elsevier ScienceDirect, JSTOR, or arXiv, the publisher page provides only a tiny sliver of the truth. Crucial context whether an open-access pre-print exists, who cited what, what real implementations are hosted on GitHub, and what prior art laid the foundation is scattered across dozens of disconnected tools and databases.

Information is rarely missing; instead, the workflow is prohibitively expensive in human attention. Researchers spend ten to twenty minutes per paper switching browser tabs, searching metadata registries, and cross-referencing authors just to answer a fundamental question: “Is this paper worth twenty minutes of deep reading?” We built Archyve to answer that question in under thirty seconds.

02 / Workflow Evolution

Without archyve vs. With archyve

Without ArchyveManual & Friction-Heavy
  1. 1.Encounter paper on paywalled publisher portal.
  2. 2.Search DOI manually across Google Scholar and Crossref.
  3. 3.Search Unpaywall or arXiv repositories for legal Open-Access PDFs.
  4. 4.Related work and research context are spread across references, papers, and separate tools.
  5. 5.Paste raw unstructured abstracts into generic LLMs and hope for no hallucinations.
With ArchyveSingle Flow
  1. Research URL → Triggered via shortcut or browser extension.
  2. Identity Validation → Strict title & author similarity verification.
  3. Scholarly Enrichment → Automated OpenAlex, Crossref & Unpaywall sync.
  4. Evidence Retrieval → Single aggregated dashboard for all your research papers.
  5. Structured Dossier → Clear recommendation score, summaries, and context.
03 / Core Engineering

What makes archyve different

Publisher-Aware Extraction

Dedicated parsers for IEEE Xplore, ScienceDirect, Springer Nature, JSTOR, and arXiv extract exact document IDs and DOIs straight from the DOM.

Pre-Synthesis Validation

Before sending a prompt to an AI model, archyve validates external metadata against the publisher baseline to prevent hallucinating incorrect literature.

Evidence-Aware Retrieval

Related papers and GitHub implementations are retrieved from live indexes first and fed as grounded context to the model, eliminating synthetic paper names.

BYOK Local Key Vault

Bring Your Own Key architecture. Your API keys remain strictly in private to yourself, and are never saved to our database.

04 / Integrity

Built for trust, not just summaries

Most AI research tools fail silently by producing strong worded summaries of papers that do not exist or mismatching author attributions. In academic research, an ungrounded summary is worse than no summary at all.

archyve treats identity verification as a first-class engineering invariant. If an external API returns a paper with a divergent title or conflicting author list, the system flags the mismatch, computes Jaccard word-overlap similarity, and prevents the AI synthesis pipeline from proceeding until identity is confirmed. Every claim in the dossier links directly back to verified source citations.

Built With the Ecosystem

Grateful to our infrastructure & data providers.

archyve stands on the shoulders of open scientific infrastructure and powerful data platforms. We would specifically like to thank:

  • Bright DataPowering resilient web data collection and dynamic publisher DOM extraction.
  • Crossref & OpenAlexOpen scholarly metadata registries providing universal DOI resolution and citation graphs.
  • UnpaywallLocating legal, open-access full-text PDFs across institutional repositories.
  • GitHub APIConnecting theoretical scientific proposals directly to open-source code repositories.
05 / Architecture

What we built in v2

01
Supported with Modern Architecture: A complete rewrite with modern framework.
02
Caching Pipeline: Free Public-access dossiers which are pre-generated and no user login required; users can check those anonymously.
03
Browser Extension & Shortcut Bridge: Instant analysis trigger via Ctrl+Shift+H or the extension popup.
04
Multi-Provider AI Engine: Seamless switching between Google Gemini and Groq models with dynamic real-time model resolution.
06 / Roadmap

What's next

v2 represents our baseline foundation. Looking ahead, we are actively experimenting with several research directions:

  • Deeper Full-Text Parsing: Extracting figures, equations, and benchmark tables directly from open-access PDFs.
  • Citation Graph Exploration: Visualizing how papers branch from seminal foundational works over time.
  • Broader Publisher Adapters: Expanding adapter support to Nature, Science, ACM Digital Library, and PubMed.

Start researching with Archyve.

Analyze your first paper or check your settings to configure your personal API key.