发布日期:2026-08-18
收录条目:20
1. MiniMax Releases MiniMax-Music3: An Open-Weights Music Model Generating Complete Five-Minute Songs From Lyrics and a Structured Caption
- 来源:MarkTechPost
- 发布时间:2026-08-17 18:36 UTC
- 链接:https://www.marktechpost.com/2026/08/17/minimax-releases-minimax-music3/
摘要:MiniMax released MiniMax-Music3, an open-weights text-to-music model. Given lyrics with section tags and a structured caption, it generates a complete song of up to five minutes in a single pass, as 32 kHz, 16-bit stereo
2. NVIDIA Nemotron 3.5 Lightning now available in Amazon SageMaker JumpStart
- 来源:AWS ML Blog
- 发布时间:2026-08-17 18:06 UTC
- 链接:https://aws.amazon.com/blogs/machine-learning/nvidia-nemotron-3-5-lightning-now-available-in-amazon-sagemaker-jumpstart/
摘要:NVIDIA Nemotron 3.5 Lightning, an open model built for high-volume agentic workloads, is now available in Amazon SageMaker JumpStart. This post shows how to deploy the 30B Mixture-of-Experts model (3B active), which deli
3. Developing an End-to-End Document Intelligence Pipeline with docTR for OCR, Layout Analysis, KIE, Benchmarking, and Searchable PDFs
- 来源:MarkTechPost
- 发布时间:2026-08-17 17:52 UTC
- 链接:https://www.marktechpost.com/2026/08/17/end-to-end-document-intelligence-pipeline-with-doctr-for-ocr/
摘要:Develop a complete document intelligence pipeline with docTR, integrating OCR, layout analysis, and KIE for production-oriented extraction and searchable PDF creation. The post Developing an End-to-End Document Intellige
4. Build OpenClaw agents that transact with Amazon Bedrock AgentCore payments
- 来源:AWS ML Blog
- 发布时间:2026-08-17 16:19 UTC
- 链接:https://aws.amazon.com/blogs/machine-learning/build-openclaw-agents-that-transact-with-amazon-bedrock-agentcore-payments/
摘要:Give an autonomous agent a wallet and spending guardrails so it can pay for paywalled APIs, MCP servers, and web content. This post connects OpenClaw to Amazon Bedrock AgentCore payments and the x402 protocol, using the
5. Whisker’s AI-powered litter robot thinks my cats swapped bodies
- 来源:The Verge AI
- 发布时间:2026-08-17 11:00 UTC
- 链接:https://www.theverge.com/tech/978323/whisker-litter-robot-5-pro-review
摘要:The greatest invention in pet tech in recent years is the litter robot. A machine that scoops your kitties' poop so you don't have to - what else could a cat owner possibly want? How about insights into your kitty's litt
6. Anthropic explains how Claude’s invisible text watermarks will work
- 来源:The Verge AI
- 发布时间:2026-08-17 10:57 UTC
- 链接:https://www.theverge.com/ai-artificial-intelligence/980869/anthropic-claude-watermarks-synthid-text-system
摘要:Anthropic has clarified how it's planning to apply invisible watermarks to Claude-generated text in order to comply with Europe's AI transparency rules. On Friday, Anthropic announced that Claude's text marking system is
7. DeepSeek AI Releases DeepSeek Harness in Developer Preview: An MIT-Licensed Agent Harness Where Everything is a Plugin
- 来源:MarkTechPost
- 发布时间:2026-08-17 09:06 UTC
- 链接:https://www.marktechpost.com/2026/08/17/deepseek-ai-releases-deepseek-harness-in-developer-preview/
摘要:DeepSeek Harness v0.1 is an MIT-licensed agent harness where every capability is a Cordis plugin. Four runtime modes, append-only session logs, and provider-agnostic model routing. The post DeepSeek AI Releases DeepSeek
8. The Defender’s Window
- 来源:OpenAI News
- 发布时间:2026-08-17 05:30 UTC
- 链接:https://openai.com/index/the-defenders-window
摘要:AI is reshaping cybersecurity for attackers and defenders alike. Learn how OpenAI is strengthening its defenses and what security teams can do now.
9. OpenAI joins PORTS-Pike project
- 来源:OpenAI News
- 发布时间:2026-08-17 05:00 UTC
- 链接:https://openai.com/index/openai-joins-ports-pike-project
摘要:OpenAI joins PORTS-Pike project, expanding community investment and supporting thousands of Southern Ohio jobs
10. Inducing Reward-Free Judging Rubrics that Reduce Over-Crediting in Agent Evaluation
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13564
摘要:arXiv:2608.13564v1 Announce Type: new Abstract: Evaluating language-model agents at scale increasingly relies on a second language model as an automatic judge, because the gold signal, an executable environment reward, i
11. Depth-Aware Sensitivity Analysis of Mixture-of-Experts Models via Magnitude-Based Expert Masking
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13565
摘要:arXiv:2608.13565v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures scale large language models (LLMs) while preserving computational efficiency through sparse activation. Despite their widespread adop
12. Modular Cognitive Architecture Emerges in Large Language Models
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13567
摘要:arXiv:2608.13567v1 Announce Type: new Abstract: The human brain exhibits a striking degree of functional specialization, with distinct networks supporting language, formal reasoning, reasoning about other minds, and reas
13. A Year in LLM Serving: Workload Evolution, Caching and Load-Balancing
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13573
摘要:arXiv:2608.13573v1 Announce Type: new Abstract: Large Language Model (LLM) serving has become a critical cloud workload, and realistic traces are essential for motivating and benchmarking serving systems. However, existi
14. Agentao: A Governed Local-First Runtime for Tool-Using LLM Agents
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13574
摘要:arXiv:2608.13574v1 Announce Type: new Abstract: LLM agents increasingly operate as execution systems that invoke tools, modify local state, use persistent memory, and interact with external protocols. These capabilities
15. AI Evaluation Should Work With Humans
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13577
摘要:arXiv:2608.13577v1 Announce Type: new Abstract: This position paper argues that the dominant paradigm of AI evaluation (which focuses on superhuman autonomous performance and so implicitly targets the goal of replacing h
16. Stable Miscalibration in Large Language Models: A Practical View of High-Confidence Errors
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13591
摘要:arXiv:2608.13591v1 Announce Type: new Abstract: High-confidence errors in large language models are often treated as evidence of fragile internal inference. We study a different possibility: stable miscalibration, where
17. Measuring Cross-Task Behavioral Consistency in Language Model Agents
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13598
摘要:arXiv:2608.13598v1 Announce Type: new Abstract: Agent evaluation relies almost entirely on outcome metrics such as success rate, which capture whether an agent succeeds but not how consistently it behaves. We argue that
18. Cross-Disciplinary Taxonomy and Modeling of Misunderstanding Generation, Amplification, and Detection, from Pragmatics to AI Agents
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13604
摘要:arXiv:2608.13604v1 Announce Type: new Abstract: Detection of misunderstanding is an urgent problem to solve because communication has moved away from real-time, in-person interaction and is increasingly handled by AI-med
19. Active Perception for Embodied Disambiguation
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13605
摘要:arXiv:2608.13605v1 Announce Type: new Abstract: Natural language provides robots with a flexible task interface, but target ambiguity in embodied environments arises not only from user intent; it can also result from mis
20. MobileMem: Learning from a Year of Mobile Experiences
- 来源:arXiv cs.AI
- 发布时间:2026-08-17 04:00 UTC
- 链接:https://arxiv.org/abs/2608.13606
摘要:arXiv:2608.13606v1 Announce Type: new Abstract: The next generation of AI agents is increasingly moving beyond systems that answer isolated questions toward persistent personal assistants that can understand, remember, a