惯性聚合 高效追踪和阅读你感兴趣的博客、新闻、科技资讯
阅读原文 在惯性聚合中打开

推荐订阅源

Vercel News
Vercel News
Y
Y Combinator Blog
H
Hackread – Cybersecurity News, Data Breaches, AI and More
The GitHub Blog
The GitHub Blog
N
Netflix TechBlog - Medium
MyScale Blog
MyScale Blog
F
Fortinet All Blogs
Microsoft Azure Blog
Microsoft Azure Blog
H
Help Net Security
C
Check Point Blog
博客园 - 聂微东
云风的 BLOG
云风的 BLOG
M
MIT News - Artificial intelligence
U
Unit 42
WordPress大学
WordPress大学
B
Blog
Last Week in AI
Last Week in AI
人人都是产品经理
人人都是产品经理
T
Tailwind CSS Blog
D
DataBreaches.Net
G
Google Developers Blog
T
The Blog of Author Tim Ferriss
Hugging Face - Blog
Hugging Face - Blog
IT之家
IT之家

DEV Community

Authentication Security Deep Dive: From Brute Force to Salted Hashing (With Java Examples) Why AI Systems Don’t Fail — They Drift Spilling beans for how i learn for exam😁"Reinforcement Learning Cheat Sheet" I Replaced Chrome with Safari for AI Browser Automation. Here's What Broke (and What Finally Worked) How Python Borrows Other People's Work The $40 Architecture: Processing 1 Billion API Requests with 99.99% Uptime Vibe Coding: A Workflow Guide (From Zero to SaaS) Most webhook security guides protect the wrong side. The scary part is delivery. Headless CMS for TanStack Start: Build a Blog with Cosmic EU Age Verification App "Hacked in 2 Minutes" — What Actually Happened Comfy Cloud’s delete function does not actually remove files Running AI Models on GPU Cloud Servers: A Beginner Guide Event-driven media intelligence with AWS Step Functions and Bedrock I scored 500 AI prompts across 8 quality dimensions — here's what broke How to Call Google Gemini API from Next.js (Free Tier, No Backend Needed) The Portal Protocol: Reclaiming Human Connection in the Age of AI How to Fix Your Team's Scattered Knowledge Problem With a Self-Hosted Forum Intro to tc Cloud Functors: A Graph-First Mental Model for the Modern Cloud Designing Multi-Tenant Backends With Both Ownership and Team Access I Built a Neumorphic CSS Library with 77+ Components — Here's What I Learned PostgreSQL Performance Optimization: Why Connection Pooling Is Critical at Scale Cómo construí un SaaS multi-rubro para gestionar expensas en Argentina con FastAPI + Vue 3 🚀 I Built an Ethical Hacking Scanner Tool – Open Source Project I Replaced /usage and /context in Claude Code With a Single Statusline A Pythonic Way to Handle Emails (IMAP/SMTP) with Auto-Discovery and AI-Ready Design I Collected 8.9 Million Polymarket Price Points — Here's What I Found About How Markets Really Move EcoTrack AI — Carbon Footprint Tracker & Dashboard Everyone's Using AI. No One Agrees How. 5 self-hosted ebook managers worth trying in 2026 Building Your First AI Agent with LangChain: From Chatbot to Autonomous Assistant
Released larkos 0.3
Okerew · 2026-06-13 · via DEV Community
Cover image for Released larkos 0.3

Okerew

Larkos 0.3: GAT neuron reasoning, temporal encoder, refactored fusion head

Core architecture changes:

Add _NeuronGraphReasoner: two-layer GAT over the live neuron graph,
producing per-neuron token embeddings (MAX_NEURONS x FUSE_GRAPH_DMODEL).
Node features include state, output, layer one-hot, connection degree,
state velocity, output magnitude, and mean edge weight (D_NODE=8).
Learned per-neuron embedding ensures distinct tokens for symmetric nodes.
Add _GATLayer: hand-rolled multi-head GAT with edge-weight-modulated
scores (tanh-squashed gain), dense [N,N] adjacency mask, and self-loops.
Add _TemporalAttentionEncoder: two-layer transformer over the
[TEMPORAL_WINDOW, FOURIER_OUT_DIM] input history with learned positional
embedding, replacing flat concatenation of window frames.
Refactor _FusionTransformerHead: attends over MAX_NEURONS+3 token
sequence (GAT tokens + band_q + band_m + driver). Token-type embeddings
(4 types) distinguish token kinds. Replaces mean-pool with learned-query
attention pool (single query, softmax over sequence). Head capacity
increased: 3 layers, d_model=64, dim_ff=128.
C-side fusion (fusion_mechanism.c):

Remove BAND_N and the neuron_flat projection pipeline; neuron reasoning
is now handled end-to-end by the Python-side GAT.
BAND_Q=32, BAND_M=32, FUSION_DIM=BAND_Q+BAND_M=64.
MEM_TOP_K: 8→32, MAX_MEM_ENTRIES: 300→1200.
Training loop:

Freeze cache extended to cover driver embedding (_cached_driver) and
GAT inputs (_cached_graph_inputs). All three caches invalidated together
on target refresh. GAT runs forward_from_inputs in-graph every step
(pinned inputs, live gradient).
x_temporal detached before MAML inner loop to prevent double-backward
through the temporal encoder graph.
graph_reasoner and temporal_encoder added to optimizer and checkpoint.
Verifier re-runs temporal_encoder on cached raw sequence to avoid
reusing a consumed autograd graph.
LR sensitivity check uses relative threshold (15% of current loss)
instead of fixed absolute delta.
Runner:

Add _advance_backend(): runs C-side decision/context/neuron/attractor/
affective updates before each step() so multi-step inference sees
evolving state.
alpha and mem_weight_ratio derived from live backend context by default,
matching the training loop's per-epoch derivation.
Temporal encoder and graph reasoner included in forward path.
Checkpoint:

Saves/loads graph_reasoner and temporal_encoder (strict=False for
backward compatibility with pre-0.3 checkpoints).
FUSION_DIM and fused_cog_norm dimension mismatch detection with safe
fallback to fresh init.
cached_driver persisted alongside cached_fused_cog.

https://github.com/Okerew/larkos_models