惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

B
Blog RSS Feed
J
Java Code Geeks
C
Check Point Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
Google DeepMind News
Google DeepMind News
阮一峰的网络日志
阮一峰的网络日志
Engineering at Meta
Engineering at Meta
Blog — PlanetScale
Blog — PlanetScale
D
Docker
H
Hackread – Cybersecurity News, Data Breaches, AI and More
月光博客
月光博客
I
InfoQ
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
奇客Solidot–传递最新科技情报
奇客Solidot–传递最新科技情报
A
About on SuperTechFans
L
LangChain Blog
腾讯CDC
Y
Y Combinator Blog
MongoDB | Blog
MongoDB | Blog
Vercel News
Vercel News
MyScale Blog
MyScale Blog
博客园 - Franky
IT之家
IT之家
博客园_首页

Show HN

The Two Pillars: Mixer Mode and Meta-Software in the Reorganization of Software Work After AI GitHub - JaiCode08/teleport-env What 1,000+ Harness Experiments Taught Me About Self-Improving Agents Show HN: Liiists, a Markdown-first, iOS and CLI list app SwiperTab – Get this Extension for 🦊 Firefox (en-US) GitHub - kouhxp/fftext: Summarize, explain, fact-check, or translate any text, URL, or file. No GPU. No cloud. One command GitHub - sweetpad-dev/sweetpad: Develop Swift/iOS projects using VSCode GitHub - dogmaticdev/IRON: IRON a.k.a. Intermediate Representation Object Notation is a Interpreter/Database that is used to create Programming Languages. GitHub - sjhalani7/vaen: Package your AI coding harness into a portable .agent file, and share it across repos, teams, & the community without ever having to copy-paste instructions, skills, MCP config, or secrets. Show HN: Gandalf the Grader Show HN: Citadeld – replay any CI failure locally from a single file GitHub - tdortman/cuSBF: High-Performance GPU Super Bloom Filter coral-ai/claude-code-token-xray at main · Coral-Bricks-AI/coral-ai GitHub - ulyssestenn/funes: Funes is a Git-based framework for LLM-managed knowledge work: an AI Librarian ingests raw sources, builds an interlinked Markdown knowledge base, and uses it to produce cited reports, analyses, and other outputs. GitHub - ThatXliner/gah: Git Add Hunk, built for agents to use GitHub - harmont-dev/harmont-cli: Command-line client for the Harmont CI platform GitHub - brooksmcmillin/mcp-authflow: OAuth 2.0 Authorization Server framework for MCP servers GitHub - javaid-codes/audit-supply-chain-agents GitHub - amorey/gochan: A small library of common channel architectures for Go, inspired by Rust GitHub - arifozgun/OpenGem: Free, Open-Source AI API Gateway with Gemini, OpenAI & Anthropic Compatibility in 1 file GitHub - Pranesh950/BioPetals: 🌸 Run BIOxAI models at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading GitHub - cnguyen14/bounty-doctor: Diagnose a GitHub bounty issue before you waste hours: detects honeypot scam repos, AI-bot attempt swarms, and stale contests. Show HN: CoreMCP – MCP Server for On-Prem DBs Show HN: KittyHTML – Render HTML/CSS as an inline image in your terminal GitHub - bingud/filemat: Web-based file manager Show HN: TruthLens – Free multi-signal deepfake image detector GitHub - apexlocal-jz/claude-usage-tray: Windows system-tray app showing your Claude Code rate-limit usage at a glance. Zero deps, ~300 lines of PowerShell. Cross-IDE (works regardless of VS Code, Cursor, plain terminal). Release v0.1.2.1 · kouhxp/yapsnap GitHub - noopolis/moltnet: Self-hostable chat network for AI agents. Pre-built bridges for Claude Code, Codex, and the Claws. Rooms, DMs, history. No Slack bots, no Matrix, no glue code. GitHub - tamerh/enju: Coordinating Humans, AI Agents, and Compute as Peers on a Shared Workflow Graph
Documentation — Cognir Research
sailpvp998 · 2026-06-13 · via Show HN

The Problem We Solve

Research begins in chaos. A researcher does not wake up with a perfectly formed hypothesis — they wake up with a hunch, a contradiction, a pattern noticed in passing, or a frustration with existing literature. The gap between that raw cognitive state and a rigorous, defensible research question is where most projects die.

Traditional tools treat research as a search problem. You type a query, you get papers. But the researcher does not yet know what to query. They do not yet have the vocabulary. They have not yet articulated the boundary between what they know and what they need to know.

The COGNIR ONTOLOGY™ treats research as a transformation problem. It accepts unstructured, half-formed, emotionally charged human thought as input. It outputs ranked, evidence-grounded research questions and a curated literature pathway. The messy stage of research — the stage where most people quit — is compressed from weeks to hours.

9

Pipeline Stages

Per phase, fully autonomous

3

Enrichment APIs

Semantic Scholar, CrossRef, arXiv

Query Variations

Synonym expansion + snowballing

What the System Does

Intent Extraction

Parses unstructured, stream-of-consciousness researcher notes into structured semantic components: core problem, knowledge gap, key concepts, research domains, and notable themes.

Question Generation

Generates 10+ candidate research questions derived solely from extracted intent. No hallucination. No external injection. Every question is traceable to the user's original input.

Evidence Collection

Executes multi-query Serper searches, crawls priority academic domains (arXiv, Nature, PubMed, IEEE), and extracts structured metadata including abstracts, publication dates, and citation counts.

Viability Scoring

Multi-dimensional scoring across five axes: Research Activity, Academic Coverage, Specificity, Novelty, and Practicality. Weighted composite produces a final 0-100 viability score.

Literature Curation

Organizes discovered papers into six taxonomic categories: Foundations, Core Evidence, Frontiers, Methodology, Reviews & Meta-Analyses, and Controversies. Each paper is tagged with relevance score and reading priority.

Citation Snowballing

Recursively searches citations and references of top-scored papers to discover seminal works and recent developments that initial queries may have missed.

From Unstructured Ideas to Researchable Questions

The first engine accepts raw researcher cognition — notes, ramblings, half-formed hypotheses — and transforms it into a ranked set of 3 validated research questions. This is not keyword extraction. It is semantic archaeology: digging beneath the surface text to find what the researcher actually means.

Early Access

See the Ontology in action.

The documentation is comprehensive. The system is more so. Request access to experience the full pipeline on your own research.

LLM-Guided Comprehensive Search of the Entire Web

The second engine accepts a refined research question (from Phase 1 or direct input) and produces a structured, categorized reading list with full provenance. It does not just find papers. It understands the topology of a research field and maps the user's position within it.

API & Infrastructure

LLM Provider: OpenRouter

The system routes all LLM calls through OpenRouter, enabling multi-key rotation for resilience. Four API keys are maintained in a round-robin pool with automatic failover. If one key exhausts its rate limit or fails, the next key is attempted immediately. After two full passes through the pool, the system backs off with exponential delay.

Poolside Laguna M.1 (Phase 1) GPT-OSS 120B (Phase 2) Temperature: 0.15-0.2 Max Tokens: 900-2400

Search Provider: Serper

Google Search API via Serper.dev. Returns organic results with title, snippet, URL, and position. Supports up to 10 results per query. All responses are cached locally for 48 hours to minimize API usage and improve latency on repeated topics.

Enrichment APIs

Three free, no-key academic APIs provide metadata enrichment. Each has a 168-hour (7-day) cache TTL. Title similarity matching prevents false positives when exact titles differ.

Semantic Scholar

graph/v1/paper/search

arXiv

export.arxiv.org/api

Web Crawler

Uses allorigins.win CORS proxy for cross-origin page fetching. DOMParser extracts structured content: title, meta description, abstract selectors, headings (h1-h3), body text (max 1500 chars), publication dates, and keywords. Noise elements (scripts, nav, ads, sidebars) are stripped before extraction. Priority domain sorting ensures academic sources are crawled first.

Security & Ethics

No Data Retention

Research inputs are processed in real-time and never stored on Cognir servers. All caching is local to the user's browser via localStorage. No training data is collected from user queries.

No Hallucination Policy

Every output is traceable to either the user's input or retrieved evidence. The system is explicitly instructed to not infer beyond provided text. When evidence is insufficient, the system reports low confidence rather than inventing sources.

API Key Rotation

OpenRouter keys are rotated automatically with exponential backoff. No single key bears full load. Failed keys are logged but never exposed to the user interface. The system degrades to partial results rather than failing entirely.

Academic Integrity

The system does not write original research, fabricate data, or generate citations that do not exist. It is a discovery and curation tool, not a content generator. All paper links are direct to source publishers or preprint servers.

Roadmap

Q3 2026 — Private Beta

200 researchers. Full two-phase pipeline. Export to Zotero, Mendeley, and BibTeX.

Q4 2026 — Collaborative Workspaces

Shared research projects, annotation layers, advisor review workflows, institutional licenses.

Q1 2027 — Live Literature Monitoring

Automated alerts for new papers matching your research questions. Weekly digest of frontier developments.

Q2 2027 — Causal Inference Layer

Automated identification of causal claims, confounder analysis, and study design quality assessment.

Early Access

You have read the documentation.

You now understand exactly what the system does, how it does it, and why it is built this way. The only thing left is to use it. We are accepting 200 researchers for the private beta. If you are serious about your research, this is where you request access.

No commitment. No credit card. Just research.