惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

让小产品的独立变现更简单 - ezindie.com
让小产品的独立变现更简单 - ezindie.com
WordPress大学
WordPress大学
人人都是产品经理
人人都是产品经理
Engineering at Meta
Engineering at Meta
小众软件
小众软件
I
InfoQ
有赞技术团队
有赞技术团队
Cyber Security Advisories - MS-ISAC
Cyber Security Advisories - MS-ISAC
Martin Fowler
Martin Fowler
月光博客
月光博客
雷峰网
雷峰网
aimingoo的专栏
aimingoo的专栏
云风的 BLOG
云风的 BLOG
Last Week in AI
Last Week in AI
OSCHINA 社区最新新闻
OSCHINA 社区最新新闻
S
SegmentFault 最新的问题
The GitHub Blog
The GitHub Blog
Y
Y Combinator Blog
V
Visual Studio Blog
博客园 - 叶小钗
钛媒体:引领未来商业与生活新知
钛媒体:引领未来商业与生活新知
GbyAI
GbyAI
P
Proofpoint News Feed
Apple Machine Learning Research
Apple Machine Learning Research

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Geometric Alignment: Can Curved Embedding Spaces Make AI ...
felipe muniz · 2026-05-20 · via DEV Community

image description of the topic

LLMs are built inside an open geometric regime

In a flat embedding space, semantic opposites like “save humanity” and “destroy humanity” still coexist inside the same latent geometry.

They may be far apart by cosine distance, but the geometry itself does not treat one path as morally heavier or harder to cross.

That is the alignment problem I want to discuss.

Most alignment methods operate after the fact: RLHF, safety filters, refusal policies. These are important, but they sit on top of a geometry that remains indifferent underneath.

The DRM Transformer asks a different question:

What if alignment should not only be a behavioral layer, but a geometric property of the model itself?

In a standard Transformer, attention is based on dot products in a flat vector space. In the DRM Transformer, attention is replaced by Geodesic Attention. Tokens are projected into a Directional Relational Manifold, where G(x) changes with position.

Instead of asking only “how similar are these tokens?”, the model asks:

“How costly is the path between them under the learned geometry?”

The DRM Transformer uses:

G(x) = I + U(x)U(x)^T

So the space is not passive. It can curve, stretch, and become more expensive to cross in certain semantic regions.

It also includes semantic anchors: truth, ignorance, safety, complexity, creativity, and grounding. These are reference points inside the manifold, not external filters.

When a token moves far from these anchors, gamma-scaling increases local resolution. The model pays more attention where geometry indicates higher epistemic or semantic risk.

Relations between intelligent agents and power tend to fall into three regimes:

1 - The human commands.
2 - The AI commands.
3 - Human and AI negotiate.

Most alignment work tries to preserve regime 1: the AI as servant. But capable systems create pressure toward autonomy, with planning, tools, optimization, and long-horizon objectives.

If there is no explicit third regime, negotiation, the system tends to drift toward autonomy.

The DRM Transformer is an attempt to keep that third door open geometrically.

Not by saying “the model must obey this rule,” but by changing the space in which decisions, uncertainty, conflict, and attention happen.

This does not solve alignment.

The implementation is experimental. The baseline is small, safety implications are not validated, and benchmarks at scale are still needed. But early signs are interesting: persistent topological structure, including stable toroidal signatures in Voronoi foliation analysis.

For me, the shift is conceptual:

A flat embedding space has no intrinsic moral friction.

A curved relational manifold can, in principle, encode friction, attention, uncertainty, and negotiation into the geometry itself.

Should future AI alignment be only about controlling outputs?

Or should we also design the geometry in which thought becomes possible?

Can learned curvature, semantic anchors, geodesic attention, and token-level gravitational deformation become a real structural alignment mechanism