惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

云风的 BLOG
云风的 BLOG
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
博客园 - 叶小钗
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
V
V2EX
酷 壳 – CoolShell
酷 壳 – CoolShell
月光博客
月光博客
人人都是产品经理
人人都是产品经理
宝玉的分享
宝玉的分享
博客园 - 司徒正美
WordPress大学
WordPress大学
Microsoft Azure Blog
Microsoft Azure Blog
罗磊的独立博客
Vercel News
Vercel News
T
The Blog of Author Tim Ferriss
T
Tailwind CSS Blog
A
About on SuperTechFans
Apple Machine Learning Research
Apple Machine Learning Research
L
LangChain Blog
让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
V
Visual Studio Blog
S
SegmentFault 最新的问题
Google DeepMind News
Google DeepMind News
博客园 - 聂微东

Hacker News - Newest: "AI"

AI can't read an investor deck AI as an attorney? Student uses ChatGPT, Gemini to sue UW over alleged racial discrimination Hacking MCP Servers in AI Systems – The Rug Pull: Tool Changes After Approval GitHub - MeepCastana/KubeezCut: Free Web based video editor Can AI judge journalism? A Thiel-backed startup says yes, even if it risks chilling whistleblowers Coming soon: 10 Things That Matter in AI Right Now DARPA built an AI to fact-check enemy weapons claims What explains heterogeneity in AI adoption? When AI Meets Muscle: Context-Aware Electrical Stimulation Promises a New Way to Guide Human Movements - Department of Computer Science AI Changed How We Build. It Did Not Change What Matters. Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… Meta spins up AI version of Mark Zuckerberg to engage with employees Code Mode: Let Your AI Write Programs, Not Just Call Tools | TanStack Blog GitHub - Delavalom/graft: Go framework for building AI agents. Type-safe tools, multi-provider (OpenAI, Anthropic, Gemini, Bedrock), zero vendor SDKs. India's TCS tops estimates, says new AI models did not dent services demand Gen Z's fading AI hype Strong feeling: we are in a folded AI reality GitHub - machinarii/total-recall-catalog: A reference catalog of latest knowledge retrieval, memory & RAG systems GitHub - mensfeld/code-on-incus: Give each AI agent its own isolated machine with root, Docker, and systemd. Active defense detects and stops threats automatically.. Quantization, LoRA, and the 8% Problem: Benchmarking Local LLMs for Production AI Iran war: We spoke to the man making Lego-style AI videos that experts say are powerful propaganda Powell, Bessent discussed Anthropic's Mythos AI cyber threat with major U.S. banks GitHub - immartian/bellamem: Persistent belief-graph memory for AI agents. Retrieves decisive context by importance — not recency, not RAG, not /compact. recursive-mode: The Repo-Native Operating System for AI Engineering After the attack on Sam Altman's home, will AI CEO's go on the offensive? The biggest advance in AI since the LLM Opus 4.6 vs GPT 5.4 One Prompt Unity World Generation Test “AI polls” are fake polls Client Challenge Can AI be a 'child of God'? Inside Anthropic's meeting with Christian leaders
GitHub - achaljhawar/1rok: Multi-LLM trading harness.
satoshiclad · 2026-05-14 · via Hacker News - Newest: "AI"

1rok leaderboard

1rok is a standalone harness for running portfolio-construction agents across OpenAI, Anthropic, Gemini, xAI, DeepSeek, GLM, and OpenRouter against the same financial tool surface. Agents query Alpaca, Yahoo Finance, FRED, and Tavily through an inline, in-process tool registry defined in this repo.

Live leaderboard tracking how each model's portfolio performs in paper trading on Alpaca (started 2026-01-20): investingbench.vercel.app.

Note

run produces an artifact. execute places orders. They are always separate commands. run never touches a broker; execute --live is the only path to real order placement.

Features

  • Inline tool registrylistTools / callTool over local handlers; one registry per pipeline run.
  • Eight tool groups — market, stock, research, technicals, options, earnings, portfolio, Tavily web search.
  • Seven LLM providers — OpenAI (GPT-5.2/5.4/5.5), Anthropic (Claude Opus 4.7 / Sonnet 4.6 / Haiku 4.5), Gemini, xAI, DeepSeek, GLM, OpenRouter — behind a single tool-calling loop.
  • Specialist agents — orchestrator, screener, fundamental, valuation, technical, sentiment, catalyst, macro, risk, constructor.
  • Two-stage pipelinerun emits a portfolio-construction JSON artifact; execute reads it and places orders via Alpaca (paper by default).
  • Provider-agnostic schemas — Zod definitions converted to each provider's tool-call format with shared retry/loop logic.

Agent Pipeline

Four stages, ten agents, one weekly run. Macro reads regime; Screener surfaces 25–30 candidates; six analysts score in parallel; Orchestrator composites; Constructor sizes trades; Alpaca executes (paper by default).

flowchart TD
    Macro["Macro Agent<br/><i>The Economist</i>"]:::entry
    Screener["Screener Agent<br/><i>The Scout</i>"]:::entry

    Sentiment["Sentiment<br/><i>Mood Reader</i>"]:::analysis
    Fundamental["Fundamental<br/><i>Accountant</i>"]:::analysis
    Valuation["Valuation<br/><i>Appraiser</i>"]:::analysis
    Catalyst["Catalyst<br/><i>Event Watcher</i>"]:::analysis
    Risk["Risk<br/><i>Risk Manager</i>"]:::analysis
    Technical["Technical<br/><i>Chart Reader</i>"]:::analysis

    Orchestrator["Orchestrator Agent<br/><i>The CIO</i>"]:::synthesis
    Constructor["Portfolio Constructor<br/><i>The Trader</i>"]:::execution
    Execute["Order Execution<br/><i>Alpaca API</i>"]:::execution

    Macro --> Screener
    Screener --> Sentiment
    Screener --> Fundamental
    Screener --> Valuation
    Screener --> Catalyst
    Screener --> Risk
    Screener --> Technical

    Sentiment --> Orchestrator
    Fundamental --> Orchestrator
    Valuation --> Orchestrator
    Catalyst --> Orchestrator
    Risk --> Orchestrator
    Technical --> Orchestrator

    Orchestrator --> Constructor
    Constructor --> Execute

    classDef entry stroke:#ff8c00,stroke-width:2px
    classDef analysis stroke:#888,stroke-width:1px
    classDef synthesis stroke:#22c55e,stroke-width:2px
    classDef execution stroke:#aaa,stroke-width:1px
Loading

Composite scoring weights: fundamental 20%, valuation 20%, risk 15% (inverted), technical 15%, catalyst 15%, sentiment 10%, macro gate 5%. Constructor caps at 8 positions, ≥85% invested, ≤40% per name.

Architecture

CLI (run | execute)
   │
   ▼
Provider ── TradingPipeline ── InlineToolRegistry (per run)
                │                       │
                ▼                       ▼
        Specialist agents ───── Tool handlers
                                        │
                                        ▼
                              src/data/services
                                        │
                                        ▼
                Alpaca · Yahoo Finance · FRED · Tavily
  1. Runner builds provider + TradingPipeline from model id.
  2. Pipeline instantiates one InlineToolRegistry per run.
  3. Agents execute through provider's tool-calling loop.
  4. Tool handlers call typed services in src/data.
  5. Services hit external APIs.

Layout

Path Role
src/data Provider clients, domain services, types
src/tools Inline tool definitions + registry (listTools / callTool)
src/harness/agents Specialist agents (orchestrator, constructor, …)
src/harness/providers Provider adapters + tool-loop
src/harness/pipeline Run orchestration
src/cli/1rok.ts CLI entrypoint (run, execute, help)

Requirements

  • Bun >= 1.1.0 — supported on macOS (x64/arm64), Linux (x64/arm64, glibc or musl), and Windows (x64/arm64).
  • API keys for the providers you intend to exercise (see Environment)

Quick Start

macOS / Linux (bash/zsh):

bun install
cp .env.example .env   # fill in keys you actually need
bun run typecheck

Windows (PowerShell):

bun install
Copy-Item .env.example .env   # fill in keys you actually need
bun run typecheck

Windows (cmd.exe):

bun install
copy .env.example .env
bun run typecheck

Run a portfolio-construction pipeline:

bun run 1rok -- run --model gpt-5.2-medium

Execute the resulting orders file (paper by default). Path separators differ per OS:

# macOS / Linux
bun run 1rok -- execute ./results/openai/gpt-5.2-medium/portfolio-construction-2026-04-16T07-00-00.json
# Windows PowerShell
bun run 1rok -- execute .\results\openai\gpt-5.2-medium\portfolio-construction-2026-04-16T07-00-00.json

Warning

--live places real orders. Without it, execution targets paper-api.alpaca.markets.

bun run 1rok -- execute ./results/<...>.json --live
bun run 1rok -- execute ./results/<...>.json --live --force

Install the CLI globally on your shell:

bun link
1rok run --model gpt-5.2-medium
1rok execute ./results/<...>.json

On Windows, bun link creates a 1rok.cmd shim on PATH; the commands above work unchanged from PowerShell or cmd.

Environment

Copy .env.example to .env. Nothing is required unless you exercise that integration.

Data providers

Var Purpose
ALPACA_API_KEY / ALPACA_SECRET_KEY Bars, quotes, positions, news, order execution
FRED_API_KEY Macro indicators, interest rates
TAVILY_API_KEY Web search, page extract, site crawl

Yahoo Finance needs no key.

LLM providers — set at least one:

OPENAI_API_KEY, ANTHROPIC_API_KEY, GEMINI_API_KEY, XAI_API_KEY, DEEPSEEK_API_KEY, GLM_API_KEY, OPENROUTER_API_KEY

Anthropic models

Model id Notes
claude-opus-4-7 Default Anthropic model
claude-opus-4-7-high Reasoning effort high
claude-opus-4-7-max Reasoning effort max
claude-sonnet-4-6
claude-haiku-4-5

The Anthropic adapter runs through @anthropic-ai/claude-agent-sdk, which ships a native claude binary as an optional dependency. The provider resolves that binary automatically for macOS (darwin-arm64, darwin-x64), Linux (linux-x64/arm64, glibc + musl), and Windows (win32-x64, win32-arm64). Override the resolved path with CLAUDE_CODE_EXECUTABLE if needed (e.g. pointing at an existing claude install, or claude.exe on Windows).

Optional

  • IROK_MODEL — default model id when --model is omitted.
  • ALPACA_API_KEY_<PROVIDER> / ALPACA_SECRET_KEY_<PROVIDER> — per-model paper credentials for execute. Falls back to the global ALPACA_* keys.
  • CLAUDE_CODE_EXECUTABLE — absolute path to the claude binary the Anthropic provider should use. Only needed if auto-resolution from the SDK's optional deps fails.

Scripts

bun run typecheck   # tsc --noEmit
bun run build       # tsc
bun run test        # bun test
bun run 1rok -- ... # CLI passthrough

Tool Groups

market, stock, research, technicals, options, earnings, portfolio, tavily. Each group registers Zod-typed definitions in src/tools/definitions/* and is aggregated through ALL_TOOLS in src/tools/definitions/index.ts.

Programmatic Use

import { createProviderFromModel } from "1rok/harness";
import { TradingPipeline } from "1rok/pipeline";

const provider = createProviderFromModel("gpt-5.2-medium");
const pipeline = new TradingPipeline({ provider });
const result = await pipeline.run();

Subpath exports: 1rok/tools, 1rok/data, 1rok/harness, 1rok/providers, 1rok/agents, 1rok/pipeline.