11 hours agoInfrastructureLaunchInfrastructureEquinix Sets Late-2026 Fabric One Beta for AI and Multicloud LinksThe planned service would put a managed orchestration layer over connections among enterprise systems, clouds and AI environments. The key test is whether it can turn broad customer requirements into dependable routing, security and recovery controls.2 min
11 hours agoInfrastructureLaunchInfrastructureDatabricks Pushes Lakebase Postgres as the Backend for Apps and AgentsThe developer push pairs managed PostgreSQL with Databricks Apps, handling application identities and connection details inside the workspace while leaving teams to choose their database environment and data-access model.3 min
Aug 31, 2026InfrastructureLaunchInfrastructureBuild American AI Takes Multimillion-Dollar Data-Center Ads to Three Election StatesThe campaign brings organized political advertising to a facilities fight that has become a flashpoint in contests for Congress and state and local government.2 min
Aug 31, 2026InfrastructureLaunchInfrastructureBroadcom Validates Five AI Model Families for VMware’s On-Premises Model ServiceThe update gives enterprises a path to run models from NVIDIA, Google, NEC, Alibaba Cloud and Z.ai inside their own infrastructure, with new sharing, routing and code-isolation services around them.2 min
Aug 29, 2026InfrastructureLaunchInfrastructureOpenRelay Routes AI Inference Across GPUs, TPUs and Trainium Through One APIThe early-access network aims to let developers buy performance rather than a particular accelerator, but its cost-savings claims and ability to deliver consistent service across mixed infrastructure remain unproven.3 min
Aug 29, 2026InfrastructureLaunchInfrastructureNvidia Rolls Out Vera Rubin as AI Competition Shifts to Data TrafficThe architecture packages compute with storage and networking hardware, betting that efficiently moving data through large AI systems is becoming its own competitive layer.2 min
Aug 29, 2026InfrastructureLaunchInfrastructureCME Targets Oct. 5 Nvidia GPU Futures Launch as CFTC Tests the BenchmarkThe proposed contracts would turn changing GPU rental rates into a financial price, not a claim on scarce hardware. Their value will hinge on a benchmark that reflects a fragmented market and draws real participation.3 min
Aug 28, 2026InfrastructureLaunchInfrastructureSageMaker Feature Store Adds 25-Record Batch Writes and a Record-Discovery APIThe new APIs replace one-record write loops and an identifier-discovery blind spot, while leaving teams to handle partial failures and non-snapshot pagination themselves.3 min
Aug 27, 2026InfrastructureLaunchInfrastructureAnthropic Puts Model-Agnostic MHS in Preview for AI-Controlled MachineryAnthropic’s proposed interface could reduce the bespoke work of connecting agents to equipment, but broad adoption still rests on future safety work and an open-source release.2 min
Aug 27, 2026InfrastructureLaunchInfrastructureDeepgram Sends SageMaker Billing and GPU Metrics to CloudWatch Without Container EgressThe release makes vendor-level cost and capacity signals visible inside AWS tools, but teams still need different metric streams for billing, feature usage and an individual endpoint’s GPU headroom.3 min
Aug 27, 2026InfrastructureLaunchInfrastructureLeiolai Launches Device-Run AI With 11M-Token Context and Output From $0.02 Per MillionThe app and API depend on users supplying computation, making coordination, latency and trust central tests of its alternative to centralized AI infrastructure.3 min
Aug 26, 2026InfrastructureLaunchInfrastructureEstuary Makes Rust Runtime Default, With Exactly-Once Delivery Still ConditionalThe change is built to keep related database updates together as they reach agents and operational systems. But a complete end-to-end delivery guarantee still relies on the receiving destination and connector.3 min
Aug 26, 2026InfrastructureLaunchInfrastructureAWS Publishes Two AgentCore Paths to Query Cross-Account Knowledge BasesThe useful constraint is not simply access across AWS accounts: generated answers require an assumed role, while many workloads may not need an agent at all.3 min
Aug 25, 2026InfrastructureLaunchInfrastructureOracle Adds AMD GPU Operator to OKE for Broader GPU Lifecycle ManagementThe optional enhanced-cluster add-on broadens Oracle’s earlier device-plugin support into driver, health and configuration management, while leaving teams responsible for supported stack combinations and component sizing.3 min
Aug 24, 2026InfrastructureLaunchInfrastructureNVIDIA Says MaxLPS Can Add Up to 40% More Rubin GPUs Within a Fixed Power BudgetThe suite treats unused rack headroom, workload tuning and cooling power as capacity to recover. Its biggest capacity claim is still a projection, while its control software remains in preview.2 min
Aug 24, 2026InfrastructureLaunchInfrastructureAWS Brings Ray Into SageMaker HyperPod With Recovery Tools and Tiered Cache on EKSThe launch moves several operational tasks behind SageMaker Studio while tying Ray teams to a HyperPod-on-EKS setup and its required add-ons.3 min
Aug 24, 2026InfrastructureLaunchInfrastructureNVIDIA Adds BlueField-4 Scale-In for 800 Gb/s AI-Factory Infrastructure OffloadThe new layer targets a growing bottleneck around GPU clusters: moving, securing and governing data without consuming the host resources assigned to AI workloads. Its performance case remains NVIDIA-supplied.3 min
Aug 24, 2026InfrastructureLaunchInfrastructureNvidia Puts Groq 3 LPX Into Production for Faster AI-Agent ResponsesThe first Nebius deployment will test Nvidia’s argument that specialized token generation can make long, tool-using AI sessions more responsive while leaving GPUs at the center of the system.3 min