aifollow.news

All AI updates

Oct 3 Oct 3 Saturday

Saturday · 1 items
IT之家 科技新闻 AI 辅助摘要 机器翻译
编辑优先级 40/100
OpenAI publishes GPT‑6 series guide with task-based model advice

The guide recommends GPT‑6 Astra for demanding reasoning, GPT‑6.1 Sol for complex coding, research and computer-use tasks, and GPT‑6 Luna for well-defined repetitive work. It advises users to specify goals, audience and constraints in prompts, and says everyday questions do not require the highest reasoning effort.

Sep 30 Sep 30 Wednesday

Wednesday · 4 items
AWS 机器学习博客 AI 辅助摘要 精选 机器翻译
编辑优先级 67/100
Amazon Bedrock AgentCore Runtime Instances supports multiple agents on one GPU instance

AWS describes a three-agent music production workflow on Runtime Instances: agents use the same session ID to share an instance and filesystem, then generate, process, and check audio. Sessions can persist for up to 14 days, but resuming with the persistent volumes depends on landing in the same Availability Zone.


推荐理由:同一会话让独立部署的智能体共享 GPU 和文件,适合需要跨天处理大文件的工作流;可用区限制也影响恢复能力。
AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 40/100
AWS publishes a prompt engineering guide for Amazon Quick

The AWS guide explains how specificity, business context, and examples can improve prompts in Amazon Quick, and introduces the CRISPE framework for complex requests. It uses RFI questionnaire processing as an example of extracting questions from inconsistently formatted Excel workbooks into structured CSV; component-specific techniques are reserved for Part 2.

AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
AWS outlines a dual-model architecture for contract analysis and structured queries

AWS outlines an architecture that uses Claude Sonnet 4.6 to extract fields from contract PDFs, Claude Haiku 4.5 to verify them independently, and a database to support queries across contracts. When the models disagree on signature status, the system calls Amazon Textract. The authors say their accuracy test covered only 20 contracts and results will vary by contract format and field complexity.

Sep 29 Sep 29 Tuesday

Tuesday · 2 items
AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
AWS publishes a tutorial for streaming speech with vLLM-Omni on SageMaker AI

AWS provides a deployment example that runs Qwen3-TTS in a vLLM-Omni container, sends text over a SageMaker AI bidirectional connection, and receives speech chunks before the full response is generated. The sample includes a Gradio client; endpoint deployment requires instance quota, and a running GPU endpoint continues to incur charges.

AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
AWS publishes SageMaker AI image-to-video deployment tutorial

AWS shows how to deploy two endpoints from the same vLLM-Omni container: FLUX.2-klein-4B generates an image through real-time inference, then Wan2.1-VACE-1.3B generates a video asynchronously and stores the MP4 in Amazon S3. The sample includes a command-line workflow and an optional Streamlit interface.

Sep 28 Sep 28 Monday

Monday · 1 items
AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
AWS publishes guide to managing Amazon Textract adapters across accounts

AWS outlines a process for promoting Amazon Textract adapters from training to production, with CloudFormation and Terraform templates and a document pre-classification pattern. Cross-account copies still require an AWS Support ticket and transfer only trained model weights; storing adapter IDs in Parameter Store lets teams update production references without redeploying the application.

Sep 26 Sep 26 Saturday

Saturday · 2 items
AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
Amazon SageMaker JumpStart offers real-time voice cloning deployment for Qwen3-TTS Base

Developers can now deploy Qwen3-TTS-12Hz-1.7B-Base from SageMaker JumpStart to a managed real-time inference endpoint and clone a voice using a few seconds of reference audio and its transcript. The walkthrough uses one 24 GB NVIDIA L4 GPU and requires a speech route when invoking the endpoint. The model also supports generating speech in another language from a reference recording.

Sep 25 Sep 25 Friday

Friday · 2 items
Claude 开发者博客 AI 辅助摘要 机器翻译
编辑优先级 47/100
Claude Code engineer shares task-based effort testing results

Thariq Shihipar recommends low effort for quick iteration in Claude Code and higher settings for tasks that need verification or edge-case testing. In internal runs, Fable 5.1’s success count on an HTML sanitizer task rose from 1/5 at low effort to 5/5 at xhigh. The figures come from five attempts per task and are not directly comparable with the public leaderboard.

AWS 机器学习博客 AI 辅助摘要 机器翻译
编辑优先级 54/100
AWS publishes tutorial for building a multi-account AI agent with AgentCore Gateway and MCP

The tutorial presents a four-account reference implementation: business teams expose data as MCP tools in their own accounts, while an agent in a central account queries them through one Gateway. Tools return only the requested results, leaving source datasets in their owning accounts; the example uses Okta authentication and per-user authorization at the Gateway.

Sep 24 Sep 24 Thursday

Thursday · 1 items
You're all caught up for this view · 10月10日 07:04