Administrator
发布于 2026-06-28 / 13 阅读
0
0

AI 每日资讯 - 2026-06-28

发布日期:2026-06-28

收录条目:20

1. Margaret Atwood says the problem with AI is ‘garbage in, garbage out’

摘要:Maraget Atwood, the storied author of The Handmaid's Tale and The Blind Assassin, was interviewed as part of the Babell Literary and Cultural Festival in Porto, Portugal. As it usually does at these things, the issue of

2. DeepSeek Releases DSpark, a Speculative Decoding Framework That Accelerates DeepSeek-V4 Per-User Generation 60–85% Over MTP-1

摘要:DeepSeek open-sourced DSpark, a speculative decoding framework that attaches a draft module to existing DeepSeek-V4 weights. It pairs a parallel draft backbone with a lightweight Markov head to cut suffix decay, then add

3. Why is Apple asking me to pay more for Big Tech’s AI obsession?

摘要:Tim Cook recently said price increases were "unavoidable" and described the company's pricing as "unsustainable." The 16-inch MacBook Pro saw its price go up by $300. The 11-inch iPad Air went from $599 to $749. Even the

4. Meta’s Astryx Brings a CLI and MCP Server to an Open-Source React Design System Agents Can Read

摘要:Meta released Astryx, an open-source React design system built on StyleX. It pairs a CSS-variable theme cascade with a CLI and MCP server, so both engineers and AI agents build using the same API. The project is in Beta,

5. Detecting and Controlling Sycophancy with Cascading Linear Features

摘要:arXiv:2606.26155v1 Announce Type: new Abstract: Interpreting and controlling model behaviors through activation steering methods requires many pairs of contrastive samples that clearly exhibit desired or undesired behavi

6. Life After Benchmark Saturation: A Case Study of CORE-Bench

摘要:arXiv:2606.26158v1 Announce Type: new Abstract: When a benchmark's accuracy saturates, it is often retired and replaced with a more challenging version. We show that this approach privileges accuracy and misses the oppor

7. Refusal Lives Downstream of Persona in Chat Models

摘要:arXiv:2606.26161v1 Announce Type: new Abstract: Linear directions in activation space have been identified for both refusal and persona traits in instruction-tuned chat models, but the two have been studied as separate m

8. AlgoEvolve: LLM-driven Meta-evolution of Algorithmic Trading Programs

摘要:arXiv:2606.26173v1 Announce Type: new Abstract: Recent work shows that Large Language Models (LLMs) can act as semantic mutation operators for the evolutionary discovery of programs and proofs. Most current applications

9. Agentic Analysis for Agentic Infrastructure: An LLM-Powered Pipeline for Comparative Governance of DAO and Corporate AI Protocols

摘要:arXiv:2606.26203v1 Announce Type: new Abstract: As AI agent protocols proliferate, the governance structures shaping their interoperability standards remain empirically underexamined. We introduce an LLM-powered comparat

10. Knowledge-augmented Agentic AI for Mental Health Medication Information Seeking

摘要:arXiv:2606.26205v1 Announce Type: new Abstract: Patients increasingly seek medication information online, yet safety knowledge for psychiatric drugs is split between regulatory adverse-event records, which are authoritat

11. Accelerating Skill Assessment in Chess: A Drift-Diffusion-Enhanced Elo Rating System

摘要:arXiv:2606.26267v1 Announce Type: new Abstract: Rating systems such as Elo serve as the gold standard for matchmaking in competitive chess. However, they inherently suffer from response lag due to their exclusive relianc

12. Governing Actions, Not Agents: Institutional Attestation as a Governance Model for Autonomous AI Systems

摘要:arXiv:2606.26298v1 Announce Type: new Abstract: Autonomous AI agents may begin to perform consequential, irreversible actions such as clinical prescribing and production software deployment. This paper observes that huma

13. COrigami: An AI Pipeline for Co-Designing Flat-Foldable Visually Recognisable Origami

摘要:arXiv:2606.26299v1 Announce Type: new Abstract: While generative AI has achieved remarkable success in solving problems with verifiable solutions, generating physical art that satisfies both strict geometric constraints

14. The Verification Horizon: No Silver Bullet for Coding Agent Rewards

摘要:arXiv:2606.26300v1 Announce Type: new Abstract: A classical intuition holds that verifying a solution is easier than producing one. For today's coding agents, this intuition is being inverted: as foundation models develo

15. How Do Tool-Augmented LLM Agents Perform on Real-World Energy Analytics Tasks?

摘要:arXiv:2606.26346v1 Announce Type: new Abstract: Agentic benchmarks have emerged across general-purpose and domain-specific settings, including finance, coding, law, and drug discovery, yet energy-domain evaluations remai

16. What We are Missing in Multimodal LLM Evaluation?

摘要:arXiv:2606.26348v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can process diverse inputs, e.g., text, images, audio, and video, and generate textual responses. While their capabilities have adv

17. OpenFinGym: A Verifiable Multi-Task Gym Environment for Evaluating Quant Agents

摘要:arXiv:2606.26350v1 Announce Type: new Abstract: Although large language model agents are increasingly applied to quantitative-finance workflows, their evaluation remains fragmented across isolated tasks, while the financ

18. Instruction Bleed: Cross-Module Interference in Prompt-Composed Agentic Systems

摘要:arXiv:2606.26356v1 Announce Type: new Abstract: Practitioners of prompt-composed agentic systems report a recurring failure mode: editing one prompt module silently shifts the behavior of others despite no shared variabl

19. Accelerating Returns and the Qualitative Engine for Science

摘要:arXiv:2606.26359v1 Announce Type: new Abstract: Ray Kurzweil described a thesis of accelerating returns, which is the most influential narratives in discussions of technological progress. Its central claim is that advanc

20. Narration-of-Thought: Inference-Time Scaffolding for Defeasible Ethical Reasoning in Large Language Models

摘要:arXiv:2606.26366v1 Announce Type: new Abstract: Standard chain-of-thought on moral dilemmas exhibits two failure modes: stakeholder collapse (the trace names at most one party with a stake in the outcome) and uncertainty


评论