CLAUDE.md não é controle
Diferencia instrução textual de bloqueio real, peça central para agentes com guardrails.
fonte: When "Do Not" Is Not Deny: Security Rules in CLAUDE.md vs Built-In Controls
Diferencia instrução textual de bloqueio real, peça central para agentes com guardrails.
Diferencia instrução textual de bloqueio real, peça central para agentes com guardrails.
fonte: When "Do Not" Is Not Deny: Security Rules in CLAUDE.md vs Built-In Controls
Memória longa em world model interativo toca diretamente interfaces espaciais e expografia generativa.
fonte: ReWorld: An Interactive World Model with Long-Horizon Memory
Mostra injeção em memória persistente de LLM agents, risco direto para agentes com histórico e personalização.
fonte: InjecMEM: Memory Injection Attack on LLM Agent Memory Systems
Mostra a tensão entre escala, interconexão e custo de sincronização em infraestrutura de ML.
fonte: Understanding the Synchronization Tax in GPU Scale-Up Domains
Explora onde modelo, engine de inferência e host viram uma única superfície de ataque.
fonte: LLMs could control their host machines by exploiting inference engines
Mede refactors longos de repositório inteiro, próximo do uso real em Portal AYA.
fonte: SWE Refactor Bench: Can Coding Agents Complete a Long-Horizon, Whole-Repository Stack Migration?
Preservar projetos, artistas e sementes é infraestrutura cultural contra perda de plataforma.
fonte: fxhash archive - a virtual museum with 400+ of the best fxhash art
Se funcionar, reduz amostras no RL ao estimar vantagens token a token.
fonte: How to Train a Critic Stably and Efficiently
Resolve uma peça formal dos LMs contínuos, ainda distante do uso operacional.
fonte: ConvergeFlow: Language Flow with Provable Convergence to Token Embeddings
Envenenamento no treino de detecção industrial é risco real para OT e sensores.
fonte: Robustness of Anomaly Detection Models for Industrial Control Systems under Training-Time Data Contamination