惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

WordPress大学
WordPress大学
GbyAI
GbyAI
P
Proofpoint News Feed
B
Blog
MyScale Blog
MyScale Blog
V
V2EX
B
Blog RSS Feed
Microsoft Security Blog
Microsoft Security Blog
量子位
Jina AI
Jina AI
博客园 - 叶小钗
Recent Announcements
Recent Announcements
有赞技术团队
有赞技术团队
罗磊的独立博客
L
LangChain Blog
I
InfoQ
云风的 BLOG
云风的 BLOG
freeCodeCamp Programming Tutorials: Python, JavaScript, Git & More
人人都是产品经理
人人都是产品经理
小众软件
小众软件
V
Visual Studio Blog
月光博客
月光博客
The Cloudflare Blog
雷峰网
雷峰网

Hacker News - Newest: "AI"

AI can't read an investor deck AI as an attorney? Student uses ChatGPT, Gemini to sue UW over alleged racial discrimination Hacking MCP Servers in AI Systems – The Rug Pull: Tool Changes After Approval GitHub - MeepCastana/KubeezCut: Free Web based video editor Can AI judge journalism? A Thiel-backed startup says yes, even if it risks chilling whistleblowers Coming soon: 10 Things That Matter in AI Right Now DARPA built an AI to fact-check enemy weapons claims What explains heterogeneity in AI adoption? When AI Meets Muscle: Context-Aware Electrical Stimulation Promises a New Way to Guide Human Movements - Department of Computer Science AI Changed How We Build. It Did Not Change What Matters. Linux rules on using AI-generated code - Copilot is OK, but humans must take 'full responsibility for the… Meta spins up AI version of Mark Zuckerberg to engage with employees Code Mode: Let Your AI Write Programs, Not Just Call Tools | TanStack Blog GitHub - Delavalom/graft: Go framework for building AI agents. Type-safe tools, multi-provider (OpenAI, Anthropic, Gemini, Bedrock), zero vendor SDKs. India's TCS tops estimates, says new AI models did not dent services demand Gen Z's fading AI hype Strong feeling: we are in a folded AI reality GitHub - machinarii/total-recall-catalog: A reference catalog of latest knowledge retrieval, memory & RAG systems GitHub - mensfeld/code-on-incus: Give each AI agent its own isolated machine with root, Docker, and systemd. Active defense detects and stops threats automatically.. Quantization, LoRA, and the 8% Problem: Benchmarking Local LLMs for Production AI Iran war: We spoke to the man making Lego-style AI videos that experts say are powerful propaganda Powell, Bessent discussed Anthropic's Mythos AI cyber threat with major U.S. banks GitHub - immartian/bellamem: Persistent belief-graph memory for AI agents. Retrieves decisive context by importance — not recency, not RAG, not /compact. recursive-mode: The Repo-Native Operating System for AI Engineering After the attack on Sam Altman's home, will AI CEO's go on the offensive? The biggest advance in AI since the LLM Opus 4.6 vs GPT 5.4 One Prompt Unity World Generation Test “AI polls” are fake polls Client Challenge Can AI be a 'child of God'? Inside Anthropic's meeting with Christian leaders
GitHub - TurixAI/TuriX-CUA: This is the official website ...
turixai · 2026-05-19 · via Hacker News - Newest: "AI"

TuriX logo

TuriX · Desktop Actions, Driven by AI

Talk to your computer, watch it work.

English | 中文

📞 Contact & Community

Join our Discord community for support, discussions, and updates:

Join our Discord

Or contact us with email: contact@turix.ai

TuriX lets your powerful AI models take real, hands‑on actions directly on your desktop. It ships with a state‑of‑the‑art computer‑use agent (achieves 80% success rate on our OSWorld‑style Mac benchmark and 64.2% success rate on OSWorld) yet stays 100 % open‑source and cost‑free for personal & research use.

Prefer your own model? Change in config.json and go.

Table of Contents


🤖 OpenClaw Skill

Use TuriX via OpenClaw with our published ClawHub skill:
https://clawhub.ai/Tongyu-Yan/turix-cua

This repo also includes local OpenClaw skill packages in OpenCLaw_TuriX_skill/:

  • macOS package in main (SKILL.md + scripts/run_turix.sh)
  • Windows package in multi-agent-windows (SKILL.md + scripts/run_turix.ps1 + agents/openai.yaml)

For installation and permissions, follow OpenCLaw_TuriX_skill/README.md.


📰 Latest News

May 11, 2026 - Now can download TuriX SuperAgent from our official web page.

April 8, 2026 - 🚀 Introducing TuriX SuperPower 3.0.0-alpha for macOS (Apple Silicon)

This is our all-in-one productivity app that combines TuriX CUA + CLI in one workflow, and adds two new capabilities:

  • TuriX-work for everyday office execution and task orchestration
  • TuriX-code for coding, automation, and engineering tasks

From writing code to handling office tasks, you can execute with CLI precision and close the loop through GUI actions in one continuous flow.

March 16, 2026 - 🐧 Linux support is now available on branch multi-agent-linux. If you want to run TuriX on Linux (for example Ubuntu), switch to that branch first:

git checkout multi-agent-linux

March 9, 2026 - Added a new OpenClaw Flash/Fast Mode skill for macOS on branch mac_legacy. If you want to use this faster, lighter setup, switch to that branch first:

git checkout mac_legacy

March 5, 2026 - Updated the Windows OpenClaw local skill on branch multi-agent-windows with direct dispatch, safer pre-flight checks, and the new OpenCLaw_TuriX_skill/agents/openai.yaml.

Earlier updates (Jan 2026 and before) - We shipped v0.3 (DuckDuckGo, Ollama, recoverable memory compression, Skills), published the TuriX OpenClaw skill on ClawHub, upgraded the core architecture to multi-model, and rolled out major model capability improvements including Qwen3-VL support and TuriX API model upgrades.

Ready to level up? Update your config.json and start automating—happy hacking! 🎉

Stay tuned to our Discord for tips, user stories, and the next big drop.


🖼️ Demos

TuriX SuperPower App Demo

TuriX SuperPower app demo

MacOS Demo

Book a flight, hotel and uber.

TuriX macOS demo - booking

Search iPhone price, create Pages document, and send to contact

TuriX macOS demo - iPhone price search and document sharing

Generate a bar-chart in the numbers file sent by boss in discord and insert it to the right place of my powerpoint, and reply my boss.

TuriX macOS demo - excel graph to powerpoint

Windows Demo

Search video content in youtube and like it

TuriX Windows demo - video search and sharing

MCP with Claude Demo

Claude search for AI news, and call TuriX with MCP, write down the research result to a pages document and send it to contact

TuriX MCP demo - news search and sharing


✨ Key Features

Capability What it means
SOTA default model Outperforms previous open‑source agents (e.g. UI‑TARS) on success rate and speed on Mac
No app‑specific APIs If a human can click it, TuriX can too—WhatsApp, Excel, Outlook, in‑house tools…
Hot‑swappable "brains" Replace the VLM policy without touching code (config.json)
MCP‑ready Hook up Claude for Desktop or any agent via the Model Context Protocol (MCP)
Skills (markdown playbooks) Planner selects relevant skill guides (name + description), brain uses full instructions to plan each step

📊 Model Performance

Our agent achieves state-of-the-art performance on desktop automation tasks:

OSWorld Benchmark — 3rd Place on the Leaderboard (50 Steps)

TuriX scores 64.2% (229.88 / 358) on the full OSWorld benchmark, ranking 3rd overall among all submitted agents. Notably, TuriX is built and optimized for macOS, where we achieve an 80%+ success rate on our self-hosted OSWorld-style Mac benchmark. We used zero Linux training data, yet still achieve a top-3 finish on OSWorld's Linux-based environment.

TuriX OSWorld benchmark score — 64.2%

TuriX performance

For more details, check our report.

🚀 Quick‑Start (macOS 15+)

We never collect data—install, grant permissions, and hack away.

0. Windows Users: Switch to the multi-agent-windows branch for Windows-specific setup and installation instructions.

git checkout multi-agent-windows

For the updated OpenClaw Windows local skill package, see OpenCLaw_TuriX_skill/README.md in that branch.

0. Linux Users: Switch to the multi-agent-linux branch for Linux-specific setup and installation instructions.

git checkout multi-agent-linux

0. Windows Legacy Users: For the previous Windows setup, switch to the windows_legacy branch.

0. macOS Legacy Users: For the previous single-model macOS setup, switch to the mac_legacy branch.

1. Download the App

For easier usage, download the app

Or follow the manual setup below:

2. Create a Python 3.12 Environment

Firstly Clone the repository and run:

conda create -n turix_env python=3.12
conda activate turix_env        # requires conda ≥ 22.9
pip install -r requirements.txt

3. Grant macOS Permissions

3.1 Accessibility

  1. Open System Settings ▸ Privacy & Security ▸ Accessibility
  2. Click , then add Terminal and Visual Studio Code ANY IDE you use
  3. If the agent still fails, also add /usr/bin/python3

3.2 Safari Automation

  1. Safari ▸ Settings ▸ Advanced → enable Show features for web developers
  2. In the new Develop menu, enable
    • Allow Remote Automation
    • Allow JavaScript from Apple Events

Trigger the Permission Dialogs (run once per shell)

# macOS Terminal
osascript -e 'tell application "Safari" \
to do JavaScript "alert(\"Triggering accessibility request\")" in document 1'

# VS Code integrated terminal (repeat to grant VS Code)
osascript -e 'tell application "Safari" \
to do JavaScript "alert(\"Triggering accessibility request\")" in document 1'

Click "Allow" on every dialog so the agent can drive Safari.

4. Configure & Run

4.1 Edit Task Configuration

Important

Task Configuration is Critical: The quality of your task instructions directly impacts success rate. Clear, specific prompts lead to better automation results.

Edit task in examples/config.json:

{
    "agent": {
         "task": "open system settings, switch to Dark Mode"
    }
}

4.2 Edit API Configuration

Get API now with credit from our official web page. Login to our website and the key is at the bottom.

In this main (multi-agent) branch, you need to set the brain, actor, and memory models. It only supports mac for now. If you enable planning (agent.use_plan: true), you also need to set the planner model. We strongly recommand you to set the turix-actor model as the actor. The brain can be any VLMs you like, we provide qwen3.5vl in our platform. Gemini-3-pro is tested to be smartest, and Gemini-3-flash is fast and smart enough for most of the tasks.

Edit API in examples/config.json:

"brain_llm": {
      "provider": "turix",
      "model_name": "turix-brain",
      "api_key": "YOUR_API_KEY",
      "base_url": "https://turixapi.io/v1"
   },
"actor_llm": {
      "provider": "turix",
      "model_name": "turix-actor",
      "api_key": "YOUR_API_KEY",
      "base_url": "https://turixapi.io/v1"
   },
"memory_llm": {
      "provider": "turix",
      "model_name": "turix-brain",
      "api_key": "YOUR_API_KEY",
      "base_url": "https://turixapi.io/v1"
   },
"planner_llm": {
      "provider": "turix",
      "model_name": "turix-brain",
      "api_key": "YOUR_API_KEY",
      "base_url": "https://turixapi.io/v1"
   }

For a local Ollama setup, point each role to your Ollama server:

"brain_llm": {
      "provider": "ollama",
      "model_name": "llama3.2-vision",
      "base_url": "http://localhost:11434"
   },
"actor_llm": {
      "provider": "ollama",
      "model_name": "llama3.2-vision",
      "base_url": "http://localhost:11434"
   },
"memory_llm": {
      "provider": "ollama",
      "model_name": "llama3.2-vision",
      "base_url": "http://localhost:11434"
   },
"planner_llm": {
      "provider": "ollama",
      "model_name": "llama3.2-vision",
      "base_url": "http://localhost:11434"
   }

4.3 Configure Custom Models (Optional)

If you want to use other models not defined by the build_llm function in the main.py, you need to first define it, then setup the config.

main.py:

if provider == "name_you_want":
        return ChatOpenAI(
            model="gpt-4.1-mini", api_key=api_key, temperature=0.3
        )

Switch between ChatOpenAI, ChatGoogleGenerativeAI, ChatAnthropic, or ChatOllama base on your llm. Also change the model name.

4.4 Skills (Optional)

Skills are lightweight markdown playbooks stored in a single folder (default: skills/). Each skill file starts with YAML frontmatter containing name and description, followed by the instructions. The planner only sees the name + description to select relevant skills; the brain receives the full skill content to guide step goals. Skills selection requires planning (agent.use_plan: true).

Example skill file (skills/github-web-actions.md):

---
name: github-web-actions
description: Use when navigating GitHub in a browser (searching repos, starring, etc.).
---
# GitHub Web Actions
- Open GitHub, use the site search, and navigate to the repo page.
- If login is required, ask the user before proceeding.
- Confirm the Star button state before moving on.

Enable in examples/config.json:

{
  "agent": {
    "use_plan": true,
    "use_skills": true,
    "skills_dir": "skills",
    "skills_max_chars": 4000
  }
}

4.5 Start the Agent

python examples/main.py

Enjoy hands‑free computing 🎉

4.6 Resume a Terminated Task

To resume a task after an interruption, set a stable agent_id and enable resume in examples/config.json:

{
    "agent": {
         "resume": true,
         "agent_id": "my-task-001"
    }
}

Notes:

  • Use the same agent_id as the run you want to resume.
  • Keep the same task when resuming.
  • Resume only works if prior memory exists at src/agent/temp_files/<agent_id>/memory.jsonl.
  • To start fresh, set resume to false, change agent_id, or delete src/agent/temp_files/<agent_id>.

🤝 Contributing

We welcome contributions! Please read our Contributing Guide to get started.

Quick links:

For bug reports and feature requests, please open an issue.

MseeP.ai Security Assessment Badge