Administrator
发布于 2026-09-11 / 1 阅读
0
0

AI 每日资讯 - 2026-09-11

发布日期:2026-09-11

收录条目:20

1. Meet Redis LangCache: A Managed Semantic Cache That Cuts LLM API Costs by Up to 90% and Returns Cache Hits Up to 15x Faster

摘要:Production LLM applications rarely receive a question nobody has asked before. Support assistants and RAG pipelines field the same intents thousands of times a day, each phrased differently, and most stacks treat every p

2. Reduce LLM latency with prefix-aware routing on Amazon SageMaker Inference

摘要:Amazon SageMaker Inference now offers prefix-aware routing, a routing strategy that sends requests sharing the same prompt prefix to the same instance so the KV cache stays warm. In benchmarks on Llama 3.1 70B, it reduce

3. NVIDIA Details BioNeMo Inference Runtime (BioIR): 2.90x Higher Boltz-2 Folding Throughput and 58.5K Residues per GPU-Hour on 8xH100

摘要:NVIDIA has detailed BioNeMo Inference Runtime (BioIR), a Python library that accelerates biomolecular structure-prediction models on NVIDIA GPUs while staying in plain PyTorch. In a matched benchmark on 1,000 human dimer

4. Reduce inference cold starts on Amazon SageMaker HyperPod with model caching

摘要:Amazon SageMaker HyperPod now supports model caching for inference, which pre-loads model weights and container images onto cluster nodes so pods read from local NVMe storage instead of downloading over the network. Lear

5. Slack can now vibe-code interactive charts and reports inside chats

摘要:A new feature coming to Slack will allow you to build interactive reports, polls, dashboards, presentations, microsites, and other tools directly inside a chat. With Slackforce Surfaces, you can describe to Slackbot what

6. OpenAI Launches the Agents API in Public Beta, Putting the Codex Harness Behind One API Call

摘要:OpenAI has released the Agents API in public beta. It gives developers the same harness and infrastructure that run Codex. OpenAI hosts and maintains the harness. Developers run the agent’s compute in an OpenAI-managed s

7. Video and image search in Amazon Bedrock Knowledge Base using Marengo 3.0

摘要:TwelveLabs Marengo Embed 3.0 is now generally available as an embedding model in Amazon Bedrock Knowledge Bases, bringing fully managed natural language search to video, image, and audio content. This walkthrough shows h

8. Schools are catching on to Big Tech’s playbook

摘要:It's the hot new thing in tech, and it's where all the jobs are. Students who don't learn to use it fall behind. And to help them catch up in time, its creators are graciously providing the resources and curriculum for l

9. Amazon Quick is now generally available on desktop

摘要:Your teams get an AI assistant that handles real work while your data stays in your environment and your conversations stay private Today, the Amazon Quick desktop application is generally available on macOS and Windows.

10. Build an end-to-end RFI questionnaire workflow using Amazon Quick Automate

摘要:Learn how to build an end-to-end RFI questionnaire workflow with Amazon Quick Automate. Read a multi-tab RFI workbook from Amazon S3, use natural-language prompts to extract and structure the questionnaire data, refine t

11. Model-agnostic PII detection with LLMs

摘要:A configurable, model-agnostic detector that turns any large language model on Amazon Bedrock into a PII detector. Because the entities to detect live in a prompt rather than in code, one detector adapts to new entity ty

12. How a researcher uses Codex and ChatGPT to search for new antimicrobial molecules

摘要:César de la Fuente’s lab uses Codex and ChatGPT to search living and extinct genomes for antimicrobial candidates to fight drug-resistant infections.

13. 3 ways to prep for your next big race with Search

摘要:Search can help runners get race-day ready with registration alerts, tailored training plans, and more.

14. Agent Evaluation Metric for multi-turn conversations

摘要:Multi-turn agents fail in ways single-turn evaluation misses: one early mistake corrupts every later turn. This post introduces the Agent Evaluation Metric (AEM), a decomposable, turn-level way to measure agent quality,

15. How AvioBook builds turnaround insights from operational data with Amazon Bedrock AgentCore

摘要:AvioBook, a Thales Group Company, prototyped Connected Analytics on Amazon Bedrock AgentCore to turn AvioBook Connect's operational data into plain-language, evidence-based answers for airline managers and dispatchers, h

16. Universal Music is launching an AI music platform with ElevenLabs

摘要:Universal Music Group is launching a new AI-powered platform that will allow users to draw from its catalog of licensed music to create song remixes, mashups, and new takes on tracks, according to an announcement on Thur

17. Now everyone can put data to work

摘要:Meet the Data agent in ChatGPT Work. Connect company data, uncover insights, and build interactive dashboards with AI using natural language.

18. Meta’s Muse AI works and creeps me out

摘要:Meta has launched its new Muse assistant, marking the company's first real foray into AI-powered productivity tools. The company says its AI agent can "take the busywork off your plate" by helping you with online shoppin

19. Why the current tech backlash feels different

摘要:This interview has been lightly edited for length and clarity. Nick Statt: Hello and welcome to Decoder, Nilay’s show about big ideas and other problems. This is Nick Statt, senior producer. And I’m joined by our brand-n

20. Mathematicians want proof OpenAI didn’t use their work

摘要:Another researcher is challenging OpenAI about the data driving its increasingly impressive array of mathematical discoveries. Just days after a bitter row erupted over whether the company's models benefited from unpubli


评论