小红花·文摘
  • 首页
  • AI Tokens🪙
  • 排行榜🏆
  • 直播
  • FAQ
Dify.AI

Sudeep Das shares how DoorDash shifts from legacy one-shot predictions to an agentic recommendation platform. He discusses leveraging language-native consumer memory, RQ-VAE semantic IDs for...

Presentation: From Models to Agents: Building Context-Aware Consumer AI at Scale at DoorDash

InfoQ InfoQ · 2026-08-15T11:00:00Z

Pinterest has revealed the Resource Provisioner Pipeline (RPP), its own Terraform execution engine. It ensures least-privilege access and needs dual-control reviews. This is important for the...

How Pinterest Secures AWS Infrastructure at Scale with a Centralized Terraform Pipeline

InfoQ InfoQ · 2026-08-10T10:00:00Z

JioHotstar explains the distributed architecture behind its real-time ad request workflow, covering ad decisioning, waterfall tiering, pacing algorithms, latency optimization, and service...

JioHotstar Explains the Distributed Engineering Behind Personalized Ad Requests at Streaming Scale

InfoQ InfoQ · 2026-08-05T14:09:00Z

该论文评估了不同检索增强生成范式在企业级语料规模扩展下的性能。实验发现,当语料超过约1000万token时,BM25检索器优于密集向量检索和文件系统代理。文件代理在小规模下表现略好,但随着规模增大,因候选发现能力不足而失效。结合BM25候选发现的Agent方法效果最佳。图RAG方法因建库成本过高或准确率低而失败。结论是BM25负责全局候选排序,Agent负责在缩小后的范围内推理。

读论文 - BM25 Wins at Scale

Measure Zero Measure Zero · 2026-08-05T00:00:00Z

AWS released Loom, an open-source reference platform on AWS Labs for governing AI agents at scale. Built on Strands Agents and Bedrock AgentCore Runtime, it implements RFC 8693 token exchange for...

AWS Releases Loom, an Open-Source Reference Platform for Governing AI Agents at Enterprise Scale

InfoQ InfoQ · 2026-07-20T10:04:00Z

As AI costs spiral, CIOs need to manage enterprise AI demand to optimize for outcomes, not just cost.

The cost of intelligence: How CIOs can manage AI demand at scale

McKinsey Insights & Publications McKinsey Insights & Publications · 2026-07-20T00:00:00Z

By Parth Jain, Rakesh Sukumar, Yingwu Zhao, Renzo Sanchez-Silva & Nathan FisherA deep dive into the engineering challenges of building a real-time service dependency map at Netflix scale: from...

Building Service Topology at Scale: Architecture, Challenges, and Lessons Learned

Netflix TechBlog Netflix TechBlog · 2026-07-13T22:44:11Z

Atlassian details the Forge billing platform built for usage-based pricing across its cloud ecosystem. It processes large-scale usage events with correct attribution, deduplication, and...

Inside Atlassian’s Forge Billing Architecture for Distributed Usage Tracking at Scale

InfoQ InfoQ · 2026-06-20T14:21:00Z

By Amer Hesson, Marcelo Mayworm, James Mulcahy, and Brittany TruongThe Problem: Managing Assets at Netflix ScaleNetflix’s Data Platform is vast. We have millions of tables in our data warehouse...

Data Projects: Managing Data Assets at Netflix Scale

Netflix TechBlog Netflix TechBlog · 2026-06-19T23:54:00Z

Vinay Chella and Akshat Goel discuss the challenges of running traditional CDC across heterogeneous databases during peak order traffic. They explain how Debezium hit limits under high load and...

Presentation: Write-Ahead Intent Log: a Foundation for Efficient CDC at Scale

InfoQ InfoQ · 2026-06-18T13:13:00Z

Adi Polak discusses the architecture required to transition from stateless prompts to state-aware, context-rich AI agents. Drawing on 15 years in distributed systems, she shares how engineering...

Presentation: Beyond Prompting: Context Engineering and Memory Management for AI Systems at Scale

InfoQ InfoQ · 2026-06-10T12:03:00Z

Stronger, more systematic partnerships between corporates and scale-ups have the potential to enable both to thrive while helping more innovation reach commercial scale and stay based in the region.

How corporate–scale-up partnering can boost Europe’s tech competitiveness

McKinsey Insights & Publications McKinsey Insights & Publications · 2026-06-10T00:00:00Z

Dropbox has unveiled Nova, an internal platform designed to orchestrate and operationalize AI coding agents across the company's engineering workflows. By Craig Risi

Dropbox Introduces Nova, an Internal Platform for Running AI Coding Agents at Scale

InfoQ InfoQ · 2026-06-05T12:00:00Z

Shopify Staff Engineer Guilherme Carreiro discusses building and scaling highly customizable platforms. Using Shopify’s Liquid theme system as a case study, he explains how to balance extreme...

Presentation: Theme Systems at Scale: How To Build Highly Customizable Software

InfoQ InfoQ · 2026-06-01T11:30:00Z

The engineering team at Meta recently outlined how the company migrated a data ingestion platform that transfers several petabytes of MySQL social graph data daily to improve reliability and...

How Meta Rebuilt Data Ingestion for Petabyte-Scale Reliability

InfoQ InfoQ · 2026-05-30T06:01:00Z

Timing, interruption, silence and recovery shape the experience as much as words

Responses are the Easy Part: What We’ve Learned Building Real-time Voice Experiences at Scale

OpenAI OpenAI · 2026-05-28T12:00:00Z

Microsoft has introduced a new AI-driven vulnerability discovery system called MDASH, a multi-model agentic security platform designed to automate large-scale code auditing across Windows and...

Microsoft Introduces MDASH for Large-Scale AI Vulnerability Research

InfoQ InfoQ · 2026-05-25T16:30:00Z

Discord has detailed how it rebuilt its database operations around a new internal orchestration framework called the Scylla Control Plane (SCP), enabling its small infrastructure team to automate...

Discord Rebuilds Database Operations Around Automation to Manage ScyllaDB at Massive Scale

InfoQ InfoQ · 2026-05-22T12:00:00Z

Grab’s Central Data Team built a multi-agent AI system to automate repetitive engineering support tasks across its data warehouse platform. The system separates investigation and enhancement...

Designing a Multi-Agent System for Engineering Support at Scale: a Case Study from Grab

InfoQ InfoQ · 2026-05-20T14:38:00Z

OpenAI recently outlined how it adapted WebRTC for low-latency voice AI at global scale. The new architecture replaced a conventional media termination model with a relay-transceiver design better...

OpenAI Outlines WebRTC Architecture for Low-Latency Voice AI at Scale

InfoQ InfoQ · 2026-05-20T12:30:00Z
  • <<
  • <
  • 1 (current)
  • 2
  • 3
  • >
  • >>
👤 个人中心
在公众号发送验证码完成验证
登录验证
在本设备完成一次验证即可继续使用

完成下面两步后,将自动完成登录并继续当前操作。

1 关注公众号
小红花技术领袖公众号二维码
小红花技术领袖
如果当前 App 无法识别二维码,请在微信搜索并关注该公众号
2 发送验证码
在公众号对话中发送下面 4 位验证码
小红花技术领袖俱乐部
小红花·文摘:汇聚分发优质内容
小红花技术领袖俱乐部
Copyright © 2021-
粤ICP备2022094092号-1
公众号 小红花技术领袖俱乐部公众号二维码
视频号 小红花技术领袖俱乐部视频号二维码