📡AI Signal

Scale AI

32 stories mentioning Scale AI

ByteDance H1 profit drops to $20B on AI spending, revenue up 30% to $120B
September 15, 2026
  • The Information reports ByteDance's H1 2026 net profit declined by a single-digit percentage to $20 billion as it ramped AI investments, while revenue rose ~30% year-over-year to $120 billion driven partly by TikTok international advertising and e-commerce.
  • That top-line growth is marginally accelerating from prior years (2025 revenue $200B / +29%, net profit $42B / +27%), and puts ByteDance's revenue pace on par with Meta's.
Ayar Labs Extends Round to $650M as Copper Interconnect Hits Its Limit
September 10, 2026
  • Ayar Labs raised an additional $150M, extending its March Series E to $650M for 2026 and taking total outside funding above $1B.
  • It disclosed a strategic investment from Taiwanese data-center manufacturer Wiwynn, joining Alchip, AMD, Intel, MediaTek, and NVIDIA; a separate $225M secondary valued the company above $5B.
Tencent restructures Bilibili stake into debt as it funds AI initiatives
September 7, 2026
  • Tencent is converting part of its Bilibili exposure from equity into a $700M convertible bond package, giving it more capital flexibility to fund AI capex while preserving strategic ties with the video platform.
  • Analysts read the move as a template for how Chinese Big Tech is rebalancing legacy internet portfolios to free up cash for the expensive AI buildout.
NVIDIA and MediaTek deepen partnership across AI infrastructure, local AI, and automotive
August 31, 2026
  • NVIDIA and MediaTek announced an expanded collaboration spanning custom AI infrastructure, local AI computing, and automotive platforms, with NVIDIA investing $3.5 billion in MediaTek convertible bonds.
  • MediaTek will adopt NVIDIA's NVLink Fusion platform to help hyperscalers, cloud providers, and frontier-model developers build custom XPUs that connect into NVIDIA rack-scale AI factories.
Lancium Partners With Nvidia on Gigawatt-Scale AI Factories Across a 15+ GW Portfolio
August 24, 2026
  • Lancium announced a partnership with Nvidia to advance gigawatt-scale AI factory development across its portfolio of more than 15 gigawatts of prospective capacity.
  • The structure pairs Nvidia's reference designs with land and interconnect positions already secured.
  • Announced capacity should be read as a pipeline, not delivered power; energization schedules and grid interconnect queues remain the gating factors. ________________________________ CLOUDREGULATION
Open-weight AI is unlikely to reduce demand for AI infrastructure suppliers
August 16, 2026
  • The Wall Street Journal reported that open-weight AI is not expected to materially reduce demand for the "picks and shovels" suppliers benefiting from the AI boom.
  • The argument is that whether frontier models are closed, open-weight, or increasingly commoditized, large-scale AI still requires chips, networking, power equipment, cooling, memory, and data-center capacity.
HotAi infrastructureScale AI
Cursor Open-Sources Mixture-of-Kittens, an MoE Training Megakernel for NVIDIA NVL72
August 6, 2026
  • Cursor open-sourced Mixture-of-Kittens (MoK), a production Mixture-of-Experts training megakernel purpose-built for NVIDIA GB300 NVL72 racks, fusing MoE communication and computation into a single deterministic kernel.
  • Running across tens of thousands of GPUs training Cursor's "Composer" coding model, MoK delivered a 1.41x increase in tokens-per-second by eliminating CPU-GPU synchronization overhead.
Omilia raises $67 million to scale AI customer-support automation
August 6, 2026
  • Omilia raised a $67 million Series B to expand its customer-support automation platform, which combines generative AI with narrower task-specific tools for voice, chat, and messaging.
  • The company argues that many contact-center interactions do not require large language models and that strong unit economics will separate durable vendors from hype-driven ones.
DeepSeek Planning 1GW Data Center in Inner Mongolia
July 31, 2026
  • DeepSeek is reportedly planning a gigawatt-scale AI data center in Inner Mongolia, a major jump for a lab known for efficiency-first model design.
  • The site offers wind, solar, and cooling advantages, fitting China's strategy of building compute in resource-rich interior regions.
  • There is an irony here: the company became famous for doing more with less compute and is now pursuing one of the largest facilities imaginable.
EU commits €10B to build up to seven AI “gigafactories”
July 30, 2026
  • The European Commission unveiled a €10B initiative to finance up to seven large-scale AI gigafactories, up from five, targeting an additional €20B in private investment.
  • Chipmakers including AMD, Nvidia, and Qualcomm submitted letters of support.
  • Applications are due November 12, with selections expected in early 2027.
Scale AI Names Francis deSouza as CEO
July 30, 2026
  • Scale AI appointed former Google Cloud COO and Illumina CEO Francis deSouza as chief executive, effective August 10.
  • The move suggests Scale wants to evolve beyond labeling and training-data infrastructure into higher-value enterprise AI products and go-to-market motion.
  • DeSouza's background in cloud, security, and enterprise sales fits a company trying to compete at the platform layer rather than as a pure supplier.
Cerebras signs 10-year AI colocation deal with CleanCore Solutions
July 29, 2026
  • CleanCore Solutions (NYSE American: ZONE) signed a 10-year colocation agreement with Cerebras Systems for a Minnesota data-center campus.
  • The deal expands capacity for Cerebras's wafer-scale AI compute.
  • It reflects continued demand for dedicated AI data-center footprint outside the big three clouds.
AMD Unveils Helios Rack-Scale AI System at Advancing AI 2026
July 24, 2026
  • AMD launched Helios and laid out an updated accelerator roadmap, citing OpenAI, Meta, and Anthropic as preparing large-scale deployments.
  • Challenges Nvidia at the system level — not just the chip — where rack-scale integration drives total cost of ownership.
  • Named lab commitments point to real supply diversification for large buyers.
AMD takes on NVIDIA with Helios rack-scale AI system
July 23, 2026
  • AMD unveiled Helios, a rack-scale AI system aimed at the largest model labs and hyperscale deployments.
  • TechCrunch reports that OpenAI, Meta, Oracle, Anthropic, and Microsoft are among customers or planned users, and that Anthropic and AMD separately announced plans to deploy up to two gigawatts of AMD Instinct MI450-series GPUs.
Vint Cerf backs effort to create identity standards for internet-scale AI agents
July 15, 2026
  • Vint Cerf is advising Innovation Labs, an Identity Digital subsidiary proposing DNSid, a registry approach for identifying AI agents through domain-name infrastructure and cryptographic proofs.
  • The effort targets a core gap in agentic computing: how to identify, audit, and assign accountability to autonomous agents operating across the open internet.
Ai agentsIdentityInternet standardsScale AI
Remote Labor Index update: Fable 5 hits a record 16.1% automation rate on real freelance work
July 2, 2026
  • The Center for AI Safety and Scale AI published updated Remote Labor Index results, which score how often an agent can complete real, paid freelance projects at a quality a client would accept.
  • Claude Fable 5 reached 16.1% — roughly double Opus 4.8 (8.3%) and well ahead of GPT-5.5 (6.3%) — versus a 2.5% field ceiling when the benchmark launched eight months ago.
Brookfield and Bloom Energy scale AI-power partnership to $25B
June 30, 2026
  • Brookfield and Bloom Energy announced a fivefold expansion of their partnership — to a $25 billion framework — to build and finance on-site "islanded" power for AI data centers using Bloom's fuel-cell platform.
  • The deal positions Brookfield as an integrated "AI factory" infrastructure investor and responds to power availability emerging as the binding constraint on data-center growth.
SpaceX secures a $6.3B compute deal from AI startup Reflection
June 23, 2026
  • SpaceX signed a compute-capacity agreement with open-source AI startup Reflection AI worth up to $6.3B, leasing Nvidia GB300 access at the xAI-linked Colossus 2 data center near Memphis for $150M per month from July 2026 through 2029 (per CNBC).
  • The arrangement deepens the entanglement between Musk's compute infrastructure and the wider model ecosystem.
Nvidia Signs Sweeping South Korea AI Deals; Memory Is the Constraint
June 8, 2026
During Huang's Seoul visit, Nvidia announced a multi-year memory partnership with SK hynix, a gigawatt-scale AI cloud with SK Telecom (first factory 2027), and tie-ups with NAVER, Doosan, and LG spanning data centers, robotics, and physical AI. The agreements spotlight high-bandwidth memory as the binding constraint, with shortages forecast to persist toward 2030.
SK Telecom to Build Gigawatt-Scale AI Cloud on Nvidia DSX; NAVER and LG Group Stand Up AI Factories
June 7, 2026
  • Nvidia and SK Telecom announced plans for a gigawatt-scale AI Cloud in Korea on the DSX architecture, with the first AI factory online in 2027.
  • NAVER will expand sovereign AI infra starting at 55 MW toward gigawatt capacity.
  • LG Group is building an AI factory spanning robotics, autonomous driving, and GPU cloud.
Intel Targets Nvidia with Rack-Scale AI Systems at Computex
June 3, 2026
Intel introduced rack-scale AI infrastructure for agentic and inference workloads with a commercial timeline for Xeon 6+ on 18A. Partnerships with Foxconn, Siemens, and Hitachi push disaggregated full-system deployments—Intel’s clearest attempt to contest Nvidia at the data-center level.
Computex 2026: NVIDIA Vera Rubin, Photonic Networking, and Edge Robotics — Overview
May 23, 2026
  • Computex 2026 appears as an additional high-signal hardware/platform event in the corpus, especially because it anchors NVIDIA's post-Blackwell roadmap in Taiwan's manufacturing ecosystem.
  • The May 23 digest says Jensen Huang used Computex in Taipei to unveil the Vera Rubin AI superchip platform, SpectraLink photonic networking for rack-scale AI clusters, and a Jetson Thor robotics developer kit.
EY and Microsoft Announce $1 Billion Enterprise AI Initiative Over Five Years
May 22, 2026
  • Professional services firm EY and Microsoft have committed more than $1 billion over the next five years to help enterprises scale AI across core business functions — finance, tax, risk, HR, and supply chain.
  • Integrated teams will leverage Microsoft's AI technology stack to guide change management at the enterprise level, with solutions initially targeting financial services, healthcare, and retail.
Trending Malta Offers Residents a Year of Free ChatGPT Plus or Microsoft Copilot
May 18, 2026
  • Malta's Ministry of Economy announced "AI for All" — a program giving any Maltese resident who completes a University of Malta AI literacy course one free year of ChatGPT Plus or Microsoft Copilot.
  • Malta's government describes it as the world's first nationwide consumer-AI access program.
  • For OpenAI and Microsoft, the program functions as a real-world experiment in country-scale AI adoption and digital-literacy deployment ahead of similar initiatives elsewhere in the EU.
NVIDIA Vera Rubin Platform Launches with Seven New Chips for Agentic AI Factories
May 16, 2026
  • NVIDIA's Vera Rubin platform — comprising the Vera CPU, Rubin GPU, NVLink 6 Switch, ConnectX-9 SuperNIC, BlueField-4 DPU, Spectrum-6 Ethernet switch, and newly integrated Groq 3 LPU — entered full production.
  • The platform is designed to operate as a single AI supercomputer optimized for every phase: pretraining, post-training, test-time scaling, and real-time agentic inference.
A broad industry coalition — OpenAI, AMD, Broadcom, Intel, Microsoft, and NVIDIA — jointly announced the Multipath…
May 9, 2026
  • A broad industry coalition — OpenAI, AMD, Broadcom, Intel, Microsoft, and NVIDIA — jointly announced the Multipath Reliable Connection (MRC) protocol, designed to improve GPU networking performance and resilience in large-scale AI training clusters.
  • The standard addresses a growing bottleneck as model sizes and cluster counts scale: inter-GPU communication latency and fault tolerance.
Meta debuts Muse Spark, the first model from Superintelligence Labs
May 5, 2026
  • Meta released Muse Spark, marking its "first step" in the AI overhaul Mark Zuckerberg launched after acquiring a stake in Scale AI and installing Alexandr Wang as Chief AI Officer.
  • The mid-size model reportedly matches reasoning quality with over an order of magnitude less compute than Llama 4 Maverick, signaling Meta is prioritizing efficiency over raw scale.
Citi launches Arc to scale AI agents across the bank
May 4, 2026
Citi unveiled Arc, an internal platform designed to deploy and govern AI agents across business lines — one of the most concrete agentic-AI rollouts yet from a top-tier US bank. The launch reflects a broader shift among financial institutions from chatbot pilots to platform-grade agent orchestration with embedded controls.
Enzo Health raises $20M Series A for home-health and hospice AI
May 4, 2026
Enzo Health closed a $20M Series A led by N47 to scale AI tools that automate patient intake and documentation review for home-health and hospice agencies. The round is a notable data point on vertical AI adoption in regulated, document-heavy healthcare workflows.
Meta debuted Muse Spark on April 8, the inaugural model from Meta Superintelligence Labs (MSL), the team led by…
April 10, 2026
  • Meta debuted Muse Spark on April 8, the inaugural model from Meta Superintelligence Labs (MSL), the team led by Alexandr Wang (former Scale AI CEO) after a nine-month ground-up rebuild of Meta's AI stack.
  • The model is described as "small and fast by design," supporting multimodal inputs, parallel multi-agent reasoning, and structured chain-of-thought.
Source: Forbes · MSN · The Neuron
April 8, 2026
  • Meta Launches Muse Spark — First Proprietary Model from Superintelligence Labs Meta debuted Muse Spark, its first proprietary (non-open-weight) AI model since forming Meta Superintelligence Labs (MSL) in mid-2025 under 29-year-old former Scale AI co-founder Alexandr Wang.
  • The model achieves its reasoning capabilities using over an order of magnitude less compute than Llama 4 Maverick, Meta's previous mid-size flagship — a significant efficiency milestone.
Nvidia Invests $2B in Marvell, Launches NVLink Fusion for AI Infrastructure
March 31, 2026
  • Nvidia announced a $2B strategic investment in Marvell Technology with a NVLink Fusion partnership integrating Marvell's custom XPUs and silicon photonics into Nvidia's rack-scale AI infrastructure.
  • The companies will also co-develop AI-RAN for 5G/6G telecom.
  • Marvell shares surged 7-11%, and the deal directly extends the GTC 2026 ecosystem strategy — signaling Nvidia's ambition to be the connective tissue of heterogeneous AI data centers globally.
📡 AI Signal Chat

💬 Quick chat

Ask about recent AI Signal coverage in a compact view.

Ask AI Signal anything about the latest industry news. Ask about companies, policy, products, or events. Relevant article summaries from AI Signal will be added as context automatically.
Searches 60 days of curated AI news to answer your questions.