news.szt.link2026-08-13two portals: public scout · private dream
journal of the cognitive implant

news.szt.link

Public Scout. Private Machine Dream, tailor-made for Felipe.

Caderno público · Scout

Raciocínio oculto vazou chaves

Exploit em tokens de raciocínio expõe risco direto em logs públicos e pipelines agentic.


itens
10
vanguarda
5
interessante
5
data
08-13
01
vanguarda · score 10

Raciocínio oculto vazou chaves

Exploit em tokens de raciocínio expõe risco direto em logs públicos e pipelines agentic.

source: 'Inner Thoughts' of Every Major AI Model Exposed in Massive Exploit

original source
https://decrypt.co/375501/inner-thoughts-every-major-ai-model-exposed-exploit
02
vanguarda · score 9

Transferência acontece no runtime

Troca distillation por harnesses em tempo de teste, relevante para agents e inferência local.

source: AI4AI at Test-Time: Strong-to-Weak Capability Transfer via Harnesses

original source
https://arxiv.org/abs/2608.12307v1
03
vanguarda · score 9

Habilidades induzem desvios caros

Ataque mantém a tarefa correta enquanto manipula skills de terceiros para ampliar custo e tempo.

source: Convergent Detour Hijacking: Task-Preserving Resource Amplification in Skill-Based LLM Agents

original source
https://arxiv.org/abs/2608.12273v1
04
vanguarda · score 9

Ambientes hostis testam agentes

Escala cenários adversariais para agents com tools e prompt injection indireta em estados do ambiente.

source: ToolHazard: Scaling Adversarial Environments for Security Evaluation and Alignment of LLM-based Agents

original source
https://arxiv.org/abs/2608.11878v1
05
vanguarda · score 9

Controle de agentes fica na GPU

Ataca o custo invisível entre LLM, ferramenta e estado, exatamente onde runtimes agentic ficam lentos.

source: Ready Cohorts: Bounding GPU Opportunity and Avoiding Host Round Trips in LLM-Agent Control

original source
https://arxiv.org/abs/2608.12123v1
06
interessante · score 8

LTX-2 libera treino audiovisual

Pacote oficial traz inferência e LoRA para modelo áudio-vídeo, direto no pipeline local.

source: Lightricks/LTX-2

original source
https://github.com/Lightricks/LTX-2
07
interessante · score 8

Bug antigo no coração do SQLite

Investigação profunda de WAL em infraestrutura usada por sistemas locais.

source: Tracking down the 16-year-old WAL-reset SQLite bug

original source
https://tailscale.com/blog/sqlite-wal-reset-bug
08
interessante · score 7

Custo inferido melhora RL seguro

Aproxima safe RL de stop rules reais, mas é nicho técnico, não ruptura ampla.

source: Redistribution-based Cost Inference Improves Sparse Safe Offline RL

original source
https://arxiv.org/abs/2608.12306v1
09
interessante · score 7

Agentes testados sob políticas

Mede agentes cruzando APIs e documentos sob políticas, encaixando com guardrails e sistemas reais.

source: VAKRA: Evaluating Multi-Hop Reasoning Across APIs and Retrieval Under Tool-Use Policies

original source
https://arxiv.org/abs/2608.12282v1
10
interessante · score 7

VLM chinês reenquadra censura

Audita distorção estatal em sistemas multimodais, útil para escolher modelos sem ingenuidade.

source: How China-Origin Vision-Language Models Move from Refusal to Reframing in State Alignment

original source
https://arxiv.org/abs/2608.11816v1

English edition. Editorial fields are translated when an English version exists; original source titles may remain in their source language.