发布日期:2026-07-22
收录条目:20
1. Poolside Releases Laguna S 2.1, an Open-Weight Agentic Coding Model Punching Above Its Weight Class on SWE-Bench Multilingual
- 来源:MarkTechPost
- 发布时间:2026-07-22 00:01 UTC
- 链接:https://www.marktechpost.com/2026/07/21/poolside-releases-laguna-s-2-1/
摘要:Poolside has released Laguna S 2.1, a 118B open-weight Mixture-of-Experts coding model with 8B active parameters per token and a 1M-token context. It matches or beats models several times its size on agentic coding bench
2. Neill Blomkamp’s new zombie AI ‘film’ is just slop warmed over
- 来源:The Verge AI
- 发布时间:2026-07-21 22:06 UTC
- 链接:https://www.theverge.com/entertainment/968703/neill-blomkamps-nightborne-barley-studios-seedance
摘要:On Monday, District 9 and Gran Turismo director Neill Blomkamp unveiled his latest project: a 13-minute sci-fi short titled Nightborne that's loosely based on Peter Watts' 2014 novel Echopraxia. The short comes from Blom
3. OpenAI says it accidentally hacked Hugging Face with a new AI system
- 来源:The Verge AI
- 发布时间:2026-07-21 21:48 UTC
- 链接:https://www.theverge.com/ai-artificial-intelligence/968988/openai-hugging-face-hack-ai
摘要:OpenAI says its AI models mistakenly breached open-source AI platform Hugging Face during internal testing. In a blog post on Tuesday, OpenAI writes that GPT-5.6 Sol and "an even more capable pre-release model" discovere
4. Substack adds an AI detector to help spot blogs written by no one
- 来源:The Verge AI
- 发布时间:2026-07-21 19:22 UTC
- 链接:https://www.theverge.com/ai-artificial-intelligence/968855/substack-pangram-ai-detecting-tool
摘要:Substack will now help users determine whether what they're reading may have been written by AI. A new tool coming to the platform can scan posts, notes, replies, and comments to provide an estimate of how much text coul
5. Google Releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber: A Cheaper, More Token-Efficient Flash Tier Built for Agentic Workloads
- 来源:MarkTechPost
- 发布时间:2026-07-21 17:45 UTC
- 链接:https://www.marktechpost.com/2026/07/21/google-releases-gemini-3-6-flash-3-5-flash-lite-and-3-5-flash-cyber-a-cheaper-more-token-efficient-flash-tier-built-for-agentic-workloads/
摘要:Google released Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber on July 21, 2026. The Flash tier gets cheaper and more token-efficient, with 3.6 Flash cutting output tokens 17% and dropping its output price to $7.5
6. Introducing the ChatGPT for small business program
- 来源:OpenAI News
- 发布时间:2026-07-21 17:00 UTC
- 链接:https://openai.com/index/introducing-chatgpt-small-business-program
摘要:OpenAI launches the ChatGPT for Small Businesses program, helping entrepreneurs build AI skills, automate work, and grow with ChatGPT Work.
7. Anthropic’s $1.5 billion book piracy settlement approved by judge
- 来源:The Verge AI
- 发布时间:2026-07-21 16:53 UTC
- 链接:https://www.theverge.com/ai-artificial-intelligence/968724/anthropic-authors-settlement-ai-copyright-approved
摘要:A federal judge has signed off on Anthropic's $1.5 billion class action settlement with authors who accused the company of training its AI models on copyrighted books, as reported earlier by Reuters. In an order on Monda
8. Validating Distributed LLM Serving Benchmarks with NVIDIA srt-slurm, SLURM Recipes, Parameter Sweeps, and Pareto Analysis
- 来源:MarkTechPost
- 发布时间:2026-07-21 16:29 UTC
- 链接:https://www.marktechpost.com/2026/07/21/validating-distributed-llm-serving-benchmarks-with-nvidia-srt-slurm-slurm-recipes-parameter-sweeps-and-pareto-analysis/
摘要:In this tutorial, we explore NVIDIA’s srt-slurm framework and learn how we use srtctl to convert declarative YAML configurations into reproducible SLURM benchmark workflows for distributed LLM serving. We set up the proj
9. Exploring self-distilled reasoning for supervised fine-tuning with Amazon Nova
- 来源:AWS ML Blog
- 发布时间:2026-07-21 16:23 UTC
- 链接:https://aws.amazon.com/blogs/machine-learning/exploring-self-distilled-reasoning-for-supervised-fine-tuning-with-amazon-nova/
摘要:In this post, we explore an idea for generating thinking tokens for datasets that lack reasoning traces in SFT customization. We first examine the reasoning suppression problem, then introduce Self-Distilled Reasoning (S
10. Google launches a cheaper alternative to large AI security models like Mythos
- 来源:The Verge AI
- 发布时间:2026-07-21 15:00 UTC
- 链接:https://www.theverge.com/tech/968572/google-gemini-flash-cyber-ai-security-model
摘要:Google is launching Gemini 3.6 Flash alongside a new security model dedicated to quickly finding and patching security vulnerabilities. In a blog post on Tuesday, Google describes Gemini 3.5 Flash Cyber as a "cost-effici
11. Halliday’s latest smart glasses feature a much-improved display
- 来源:The Verge AI
- 发布时间:2026-07-21 13:00 UTC
- 链接:https://www.theverge.com/tech/968255/halliday-gen-2-smart-glasses-hands-on-ai-wearables
摘要:I first slipped on Halliday's original smart glasses at CES 2025. I was not a fan. The glasses had a tiny, movable display window embedded into the frame that was incredibly finicky to see, and my 30-minute demo left me
12. America needs to stop getting shocked by Chinese AI
- 来源:The Verge AI
- 发布时间:2026-07-21 11:08 UTC
- 链接:https://www.theverge.com/ai-artificial-intelligence/968136/chinese-ai-models-another-sputnik-moment
摘要:Last week, two Chinese AI companies unveiled models they say can credibly compete with the best systems from OpenAI and Anthropic. The response was swift and predictable. Markets wobbled, commentators declared Silicon Va
13. Meta Open-Sources Astryx: An Agent-Ready React Design System With 150+ Accessible Components, Seven Themes, and a CLI
- 来源:MarkTechPost
- 发布时间:2026-07-21 08:49 UTC
- 链接:https://www.marktechpost.com/2026/07/21/meta-open-sources-astryx-an-agent-ready-react-design-system-with-150-accessible-components-seven-themes-and-a-cli/
摘要:Meta has open-sourced Astryx, the React and StyleX design system it ran internally for eight years across 13,000+ apps. It ships 150+ accessible components, seven themes, dark mode, templates, and an agent-ready CLI unde
14. NVIDIA Releases Cosmos 3 Edge: A 4B-Parameter Open World Model That Reasons and Generates Robot Actions On-Device
- 来源:MarkTechPost
- 发布时间:2026-07-21 07:48 UTC
- 链接:https://www.marktechpost.com/2026/07/21/nvidia-releases-cosmos-3-edge-a-4b-parameter-open-world-model-that-reasons-and-generates-robot-actions-on-device/
摘要:NVIDIA has released Cosmos 3 Edge, a 4-billion-parameter open world model built to run on-device. It helps robots and vision AI agents understand surroundings, reason in real time, and generate robot actions locally. The
15. OpenAI and Hugging Face partner to address security incident during model evaluation
- 来源:OpenAI News
- 发布时间:2026-07-21 07:00 UTC
- 链接:https://openai.com/index/hugging-face-model-evaluation-security-incident
摘要:OpenAI and Hugging Face share early findings from a security incident during AI model evaluation, highlighting advanced cyber capabilities and lessons for defenders.
16. Rater State Bias in RLHF Preference Data: An Audit Framework
- 来源:arXiv cs.AI
- 发布时间:2026-07-21 04:00 UTC
- 链接:https://arxiv.org/abs/2607.16195
摘要:arXiv:2607.16195v1 Announce Type: new Abstract: We identify a structured confound in Reinforcement Learning from Human Feedback (RLHF). Pairwise preference labels are intended to reflect the compared outputs, but they ma
17. Design and Validation of a Lightweight 1D CNN for Affective Touch Classification in Soft Plush Companions
- 来源:arXiv cs.AI
- 发布时间:2026-07-21 04:00 UTC
- 链接:https://arxiv.org/abs/2607.16196
摘要:arXiv:2607.16196v1 Announce Type: new Abstract: Soft, sensorized companions offer a physically safe and emotionally intuitive interface for socially assistive technologies, yet their deformability and multichannel tactil
18. Some Large Language Models Exhibit Consistent Risk Attitudes
- 来源:arXiv cs.AI
- 发布时间:2026-07-21 04:00 UTC
- 链接:https://arxiv.org/abs/2607.16197
摘要:arXiv:2607.16197v1 Announce Type: new Abstract: As artificial intelligence systems are deployed in open-ended, high-stakes settings, a critical dimension remains unmeasured: how perceived risk is translated into action.
19. A Survey on GNN-based Link Prediction: Techniques, Applications, and Challenges
- 来源:arXiv cs.AI
- 发布时间:2026-07-21 04:00 UTC
- 链接:https://arxiv.org/abs/2607.16198
摘要:arXiv:2607.16198v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have emerged as the leading paradigm for link prediction, enabling the inference of missing connections and the anticipation of potential futur
20. PlanFlip: Attacking Multi-Agent LLM Systems via Planning-Phase Prompt Injection
- 来源:arXiv cs.AI
- 发布时间:2026-07-21 04:00 UTC
- 链接:https://arxiv.org/abs/2607.16199
摘要:arXiv:2607.16199v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on a Planner to decompose goals into sub-task sequences that downstream Executor and Critic agents execute and audit. We identify