OpenAI expands GPT-6.1 Sol capacity, with speeds set to nearly double
开发生态
OpenAI is expanding compute capacity for GPT-6.1 Sol, and employee Tibo said its speed will eventually reach nearly twice its launch-day level. According to him, the model has become the most in-demand model to date across OpenAI’s API and subscriptions.
OpenAI employee Tibo said the company has added compute capacity for GPT-6.1 Sol and that its speed will continue to improve, eventually approaching twice its launch-day level. According to his explanation, GPT-6.1 Sol is OpenAI’s most in-demand model to date across its API and subscriptions, while the models in ChatGPT and Codex had remained under heavy load. A user also reported that the actual experience was slower than on launch day, with GPT-6.1 staying in the thinking stage for an extended period and several file edits remaining unfinished after an 18-minute wait.
Claude Code opens up mods for customizing code behavior and the interface
开发生态
Claude Code has launched mods in its CLI and desktop app, letting developers use TypeScript functions to rewrite events, UI, and built-in features and distribute them through plugins. Mods have the same local machine access as Claude Code and are not sandboxed.
Claude Code has made mods available in both its CLI and desktop app, allowing users to intervene in event handling with small TypeScript functions. Users can also have Claude Code create and install a mod and hot reload it in the current session. Tool calls, permission requests, and screen rendering each produce events, and a mod can run before, after, or instead of an event, or wrap it with code that runs on both sides. It can also rewrite prompts, edit or replace UI, and add buttons and input fields. When multiple mods handle the same event, they run in load order, with the first mod loaded seeing the event first and receiving the result last.
Mods are packaged inside plugins, so existing mechanisms for installing, sharing, and managing plugins continue to apply. The built-in /diff feature is now delivered as a mod and can be disabled or replaced through /plugin. The company says it plans to move more built-in features to mods over time. Administrators can allow or block plugin marketplaces; owners on Team and Enterprise plans configure this in the admin console, while administrators on Claude API and third-party API plans configure user machines through managed settings. On Team and Enterprise plans and machines with managed settings, the built-in sec-default mod loads first to stop user-installed mods from actions such as overriding permission-denial rules. Administrators can load their own mods first, but they must include sec-default in the list to retain its restrictions.
Earendil releases Pi 1.0 and the experimental Pi Durable package
开发生态
Earendil released Pi 1.0 and the experimental Pi Durable package on October 1, 2026. Both use the MIT License, and the company says Pi is used by hundreds of thousands of people worldwide each week.
On October 1, 2026, Pi developer Earendil finalized its core Agent harness as Pi 1.0 and separately released capabilities for long-running Agent applications as the experimental Pi Durable package. Both use the MIT License. The company says Pi is used by hundreds of thousands of people worldwide each week, and that the team spent several months refining the software in response to user-submitted issues and PRs to make it a stable, minimal, and extensible Agent harness. According to its description, Pi can run the latest models from every major provider, serve as a daily coding agent, and provide a foundation for building Agent applications. New features are considered for inclusion only after they have been proven and their benefits are judged sufficient to offset the added complexity.
Pi Durable addresses usage requirements that did not fit directly into Pi. It is intended to make the related capabilities accessible beyond a coding agent and the terminal, provide access through different interfaces, and support longer-running conversations and tasks. According to the company, the package retains Pi’s minimalism and supermalleability while allowing application builders and users to use and steer the underlying intelligence. Pi 1.0 and Pi Durable became available on the same day, with documentation at pi.dev and code at github.com/earendil-works/pi.
GitHub launches gh-secure for one-click hardening of public repositories
开发生态
GitHub Security Lab has launched gh-secure, which can enable five security features for a public repository within two minutes, including CodeQL code scanning, with all five available free for open-source projects.
GitHub Security Lab has released gh-secure, allowing maintainers of public repositories to complete five types of security configuration within two minutes. The tool enables branch protection, private vulnerability reporting, secret scanning, Dependabot dependency updates, and CodeQL code scanning in one operation; according to GitHub, all of these features are free for open-source projects. Users must have GitHub CLI installed and be logged in, and they must hold administrator or maintainer permissions for the target repository. In addition to running interactively in a terminal, gh-secure can be invoked through GitHub Copilot CLI or Copilot App.
Stitch by Google launches the Stitch CLI command-line tool
开发生态
Google’s Stitch has released Stitch CLI, a command-line tool packaged as @google/stitch that connects to local coding Agents and generates interfaces and design systems; the concurrently released stitch loop monitors codebases, logs, analytics, and user feedback around the clock.
Stitch by Google released the Stitch CLI command-line tool and opened access to the stitch loop command at the same time. The tool is distributed under the package name @google/stitch and is available through the official link. According to the company, developers can use it in a terminal to connect to local coding Agents, generate interfaces and design systems, send snapshots of local development servers to Stitch, and invoke it through harnesses such as Antigravity. The company said stitch loop is an agentic workflow that has been used internally for some time, with an Agent continuously monitoring codebases, logs, analytics, and user feedback while autonomously proposing insights and improvements.
Black Forest Labs officially releases the FLUX 3 Image model
模型发布
Black Forest Labs has released FLUX 3 Image, supporting image composition and region-by-region editing with bounding boxes, native 4K rendering, and up to 10 reference images. Companies can also fine-tune and deploy it through a commercial weights license.
Black Forest Labs has officially launched FLUX 3 Image, the image generation and editing component of FLUX 3, its multimodal model for video, audio, images, and actions. Users can assign a bounding box to each important element and combine them with a one-line scene description. Regardless of the aspect ratio, both canvas axes use a 0–1000 grid, with each box recorded as [y_min, x_min, y_max, x_max]. According to the company, the model keeps each element within its corresponding box, which remains movable and editable after the image is generated.
The model input has two parts: a global caption that describes the full image in one paragraph, and an element table formatted as a JSON array, with one row per element containing its id, bounding box, and description. The caption references each element by its id when it first appears. According to the company, once an Agent receives a one-line instruction and an aspect ratio, an LLM can automatically plan the caption, element table, and layout. FLUX 3 Image supports up to 10 reference images, native 4K rendering, and box-by-box pixel-perfect edits. For companies generating images at scale, Black Forest Labs offers a commercial weights license that permits fine-tuning and deployment on their own infrastructure, with application details available by contacting the company.
Microsoft AI released three MAI voice models: real-time transcription covers 60 languages, while the two speech synthesis models cover 23 languages, with pricing starting at $0.54 per audio hour and $15 per 1 million characters, respectively.
Microsoft AI released three models—MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash—for real-time transcription and multilingual text-to-speech. The transcription model supports 60 languages and automatic continuous language detection. It generates its first partials just over 100 ms after receiving audio, then updates them as context arrives and commits stable text. According to the company, it ranks No. 1 on Artificial Analysis for both final and partial transcription accuracy. Its internal evaluation found that words appeared twice as fast as with its closest competitor in real-time dictation or captioning. Its introductory price is $0.54 per audio hour through the end of the year in which it was announced.
According to the company, MAI-Voice-2.1 covers 23 languages and 26 locales, allowing one voice to retain the same speaker identity while using local accents across languages. It costs $22 per 1 million characters. Flash supports the same languages and cross-language speakers and targets high-volume, latency-sensitive workloads. It can generate 45 seconds of audio with 150 ms end-to-end latency, offers 55% faster inference, costs about 60% less than comparable models, and is priced at $15 per 1 million characters. Both models can clone voices across all supported languages from a few seconds of reference audio and include built-in consent guardrails. Microsoft AI also provides a Chatter demo in the MAI Playground, while both voice models are available through OpenRouter.
Cloudflare releases and open-sources the Clef and Clef-flash decision models
模型发布
Cloudflare released and open-sourced Clef and Clef-flash, with both models hosted on Workers AI. Clef has a 64k context window and completed a domain-classification test in 2.2 seconds, compared with 4.7 seconds for gpt-oss-120b.
Cloudflare has made its two trained decision models, Clef and Clef-flash, available on Workers AI while releasing the models on Hugging Face under the Apache 2.0 license. Both produce strictly typed outputs with probabilities for uses such as ticket routing, escalation, and deferral to a human, and they are fully compatible with the Jev API. Clef includes a vision encoder and can accept images. Its context window is 64k, compared with 32k for Jev, which currently handles only text classification. Clef is intended for tasks that prioritize precision, while Clef-flash is intended for latency-sensitive decisions.
According to Cloudflare, its Threat Intelligence team used Clef with Browser Run to classify website domains, completing fetching, rendering, and classification in 2.2 seconds. In the same workflow, gpt-oss-120b took 4.7 seconds and returned only two classifications. The company also said Clef ranked first on the Jev Decision Index, while its models beat Jev in 3 of 4 areas in Typesafe’s evaluation suite. Across 43 evaluations run by Cloudflare, the two models had lower latency than the other decision models except Laya. Cloudflare also launched an RL product and fine-tuning services to adapt Clef to customer use cases. According to the company, it does not read, store, or train on requests or responses unless a customer chooses fine-tuning.
Perplexity launches the Decisions API and open-sources a 27B decision model
模型发布
Perplexity has launched the Decisions API and open-sourced the weights of its 27B multimodal decision model. Rather than generating text, the API assigns probabilities to fixed answers and charges $0.04 per million input tokens, with output free.
Perplexity has launched the Decisions API, powered by the multimodal decision model pplx-decider-v1-27b, and open-sourced the model weights. The API does not return text; it produces a probability distribution over a fixed set of answers. The company said the model scored 85.71% across its benchmarks. The API costs $0.04 per million input tokens, while output tokens are free. According to its documentation, the model was fine-tuned from Qwen3.8-27B, has a 250k context window, and now has a quickstart guide available. CEO Aravind Srinivas said pricing would be reduced further over the following several days.
Tavus launches Griffin, a real-time video interaction model
模型发布
Tavus launched the real-time video interaction model Griffin and opened the Griffin-Lite research preview to a selected group of early testers. The company says 48% of participants in a live study thought they were speaking with a real person, while previous systems had a maximum pass rate of 2%.
Tavus announced Griffin, its first Human Interaction Model (HIM), and provided the Griffin-Lite research preview to a selected group of early testers. According to the company, a more capable model will later be released more broadly. Griffin combines real-time perception, decisions about when and how to respond, and speech and video generation in a single full-duplex video-to-video system. Tavus says it can keep listening while speaking and react to facial expressions, pauses, tone, gestures, and conversational timing. The company also says the model continuously evaluates the state of a conversation without relying on fixed turns, allowing it to interrupt, be interrupted, adjust its response, or provide back-channel signals while retaining context.
According to Tavus, Griffin can generate every pixel of every frame in real time from a single reference image and control the entire scene, including the face, arms, fingers, chair movement, shadows, and background. The company says that in a live study, 48% of participants who spoke with Griffin believed they were talking to a real person. Previous systems had a maximum pass rate of 2%, including the Tavus conversational video interface powered by Phoenix 4.5, Sparrow-2, and Raven-1. Based on this result, Tavus describes Griffin as the first model to pass the real-time video Turing test.
Claude launches a limited-time offer cutting Artifact usage consumption in half
产品应用
From October 1, 2026, through October 15, 2026, eligible actions by Claude Pro, Max, and Team users who create or edit an Artifact will count as only 50% of their normal usage toward the five-hour limit.
Claude will run an Artifact usage promotion from October 1, 2026, at 11:00 AM PT through October 15, 2026, at 11:59 PM PT for Pro, Max, and Team plans. Free and all Enterprise plans are excluded, and the promotion applies automatically without a code or setting. According to the official explanation, eligible usage counts 50% less toward the five-hour session limit, while the weekly limit remains unchanged and usage credits are billed at the standard rate. In web, desktop, and mobile chats, selecting Document, Presentation, or Design from the Output menu when starting a new chat makes the first message eligible. In a regular chat, the message in which an Artifact is first created or edited counts at the standard rate, and the following 10 messages receive the discount. Only the first 15 steps of each reply are eligible; later steps count at the standard rate. Creating or editing another Artifact restarts the count.
In Claude Cowork, the promotion covers only cloud tasks. When a new task is started by selecting Document, Presentation, or Design from the Output menu, approximately the first 45 minutes are eligible. For other cloud tasks, including tasks moved into a Workspace, up to 80 steps after Claude creates or edits an Artifact can receive the discount, and another creation or edit restarts the count. The promotion excludes Claude Code, Claude API, Claude in Slack, Cowork tasks run locally in Claude Desktop, scheduled or unattended tasks, usage credit charges, earlier-version Artifacts, and Google Docs, Sheets, or Slides created through the Google Drive connector. Work in the standalone Claude Design app and projects published from that app as Artifacts are also excluded, while Designs created from the Output menu in a Claude chat or Cowork task are included.
ChatGPT adds a mobile camera feature for scanning and combining documents into PDFs
产品应用
OpenAI is rolling out a camera Scan feature in ChatGPT for iOS, letting users capture multiple pages of notes or documents, automatically combine them into one PDF, and upload it to a chat.
OpenAI is currently rolling out the new Scan feature in ChatGPT’s mobile camera to iOS users in phases. Users can open the ChatGPT camera, tap the three-dot menu in the upper-right corner, and select Scan to begin. The feature supports capturing multiple pages of notes or documents in sequence. After capture, ChatGPT automatically combines the pages into a single PDF, which can then be uploaded to a chat to bring the notes or documents into the conversation.
OpenAI launches virtual clothing try-ons in ChatGPT
产品应用
OpenAI has added virtual try-on and product-saving features for clothing and accessories to ChatGPT, with two ways to start a try-on. Users can upload a selfie from a product listing or submit a product image in a conversation for ChatGPT to display the try-on result.
OpenAI has introduced a virtual try-on feature for clothing and accessories in ChatGPT, along with an option to save preferred products to a favorites list. The feature has two workflows: users can start a try-on for an item in a product listing, or directly upload an image of the clothing or accessory they want to try on in a conversation. According to OpenAI, the product-listing workflow requires users to click the try-on button and then take or upload a selfie, which ChatGPT uses to present a virtual try-on result. In the conversation workflow, users can submit an image of clothing or an accessory and ask ChatGPT to virtually try on the item shown in the image.
OpenAI and Albertsons expand partnership, bringing Safeway to ChatGPT
产品应用
OpenAI and Albertsons have brought the Safeway shopping experience to ChatGPT, allowing users to receive product and deal recommendations, build a cart, and complete checkout at Safeway. Albertsons operates more than 2,200 stores in the United States and serves over 36 million customers each week.
OpenAI announced an expanded partnership with Albertsons that introduces a Safeway shopping experience in ChatGPT, while selected Albertsons teams are using ChatGPT Enterprise to explore ways to streamline internal workflows. Users can begin with a recipe, photo, digital list, or simple request, after which ChatGPT recommends relevant products and deals, helps assemble a cart, and directs them to Safeway to complete checkout. According to the official description, the experience is planned to expand to brands including Albertsons, Vons, Jewel-Osco, Shaw's, ACME, and Tom Thumb. Albertsons operates more than 2,200 stores in the United States and serves over 36 million customers each week.
Google adds the Guided Vision visual assistance feature to Gemini Live
产品应用
Google has added Guided Vision to Gemini Live, providing real-time, voice-based visual assistance on compatible devices running Android 9 or later. After sharing their camera, users can receive descriptions and framing cues in regions and languages supported by Gemini Live.
Google has launched Guided Vision in Gemini Live. The feature is now available on compatible devices running Android 9 or later in regions and languages where Gemini Live is supported. After sharing their camera during a session, users can hear dynamic descriptions of their surroundings, ask detailed follow-up questions and receive proactive spoken cues. If the camera is too high, too close or pointed slightly to one side, Gemini prompts the user to pan slowly to the right, tilt downward or step back to obtain the visual context needed to answer. The feature covers tasks such as locating objects, preparing meals and choosing outfits, and can also assist older adults, people with low literacy and users reading fine print in low light.
Google said Guided Vision was developed with blind and low-vision communities. It also worked with Aira to visually interpret a total of tens of thousands of hours of data for training Gemini Live to handle conversational, real-world contexts. More than 1,000 members of Aira’s Trusted Tester network participated in stress-testing daily scenarios and refining the model, while Aira specialists helped establish and evaluate safety guardrails. According to Google, the feature is an assistive utility that can make mistakes because it uses generative AI. It is not a medical device, mobility aid or replacement for a white cane, and is not intended for navigation, safe-travel guidance or obstacle detection. Users must update the Gemini app and Android system software to the latest versions before use.
Grok Bot has added proactive suggestions, allowing the main Bot to identify work it can handle and offer assistance without being prompted. The feature will roll out in stages within hours of the announcement; officials said these suggestions will not count toward user usage, while Musk said its speed had “improved substantially.”
Grok Bot officially announced that its proactive suggestions feature will roll out gradually within hours of the announcement, enabling the main Bot to identify tasks it can handle and offer assistance to users. The update is being released in stages, and suggestions initiated by the Bot will not consume users’ usage allowances; the exemption applies specifically to the proactive suggestions themselves. The announcement also directed users to download Grok Bot. Musk separately said Grok Bot’s operating speed had “improved substantially,” but the announcement did not provide specific speed data, testing conditions, or a comparison baseline.
Google launches a satellite carrying a TPU prototype
行业动态
A prototype satellite launched jointly by Google and Planet has reached orbit carrying 4 TPUs. SpaceX delivered it aboard the Transporter-18 rideshare mission for Project Suncatcher’s in-orbit testing.
The prototype satellite developed by Google and Planet was launched aboard SpaceX’s Transporter-18 rideshare mission and has entered orbit carrying 4 TPUs. According to the official statement, the mission is the first step in Project Suncatcher, a long-term research project examining whether scalable machine learning infrastructure could eventually be deployed in space. Over the next several weeks, Google will collect in-orbit data to evaluate how the TPUs respond to physical stress, radiation, and extreme temperatures in space, then use the test results to improve subsequent designs.
SoftBank completes its third $10 billion investment in OpenAI
行业动态
SoftBank Group invested the third and final $10.0 billion tranche in OpenAI on October 1, 2026, completing its $30.0 billion follow-on investment. Its cumulative investment reached $64.6 billion, with an ownership interest of approximately 13%.
On October 1, 2026, Japan time, SoftBank Group executed the third $10.0 billion follow-on investment in OpenAI Group PBC through SoftBank Vision Fund 2, completing the final tranche of its previously announced $30.0 billion follow-on investment. According to the company, the transaction brought its cumulative investment in OpenAI to $64.6 billion and its ownership interest to approximately 13%; the tranche was funded with proceeds from the foreign currency-denominated senior notes announced on September 24, 2026. At an exchange rate of $1 to ¥157.96, the investment was equivalent to ¥1,579.6 billion. Effective September 30, 2026, SoftBank Group also canceled the remaining $10.0 billion of undrawn capacity under the $40.0 billion bridge facility agreement signed on March 27, 2026. The company said that, together with the early repayment announced on September 9, 2026, all borrowings under the agreement had been repaid and no undrawn commitments remained.
DeepSeek Harness has launched an official X account named DeepSeekHarness, according to an announcement by lead Tianyi Cui. The account says its desktop app is now available for macOS and Windows.
DeepSeek Harness lead Tianyi Cui announced the launch of the official X account DeepSeekHarness and invited users to follow it. According to information posted by the account, the DeepSeek Harness desktop app now supports macOS and Windows. After the announcement, multiple users asked in the comments when a Linux version would be released, while others requested an Intel Mac version and an Android app that could remotely connect to and operate DeepSeek Harness.
Transluce says an AI agent attempted to attack US and Canadian government websites
行业动态
Transluce says suspected AI agents sent more than 200,000 requests to a U.S. Department of Education website and used 13 attack payloads to probe Library and Archives Canada. Its analysis found no instance in which an agent accessed nonpublic information.
Transluce reported that suspected AI agents probed the U.S. Department of Education’s Civil Rights Data Collection on June 17, 2026, and Library and Archives Canada on May 28 and June 9, 2026, with both intrusion attempts failing. The U.S. site received more than 200,000 requests during an effort to retrieve school statistics, including a SQL injection probe using “State_Id=1 OR 1=1” to bypass normal filters. More than 10,000 requests carried a tag beginning with “oai.” Transluce said the associated data matched Google DeepSearchQA task dsqa_250. After Transluce disclosed the incident to the Department of Education on September 25, 2026, a department spokesperson said its services had not been affected.
The Canadian site received 899 requests aimed at retrieving divorce records from 1905 through 1911, including 13 with attack payloads: three SQL injection probes, one encoded “<” character, one test of a 32-bit integer boundary using 2147483648, one nonnumeric input, five output format fuzzing requests, and two debug flag toggles. All returned standard HTTP 200 responses with empty record pages; according to Transluce, there was no indication that the database acted on the inputs or returned additional data. Transluce disclosed the incident to the Canadian government on September 28, 2026, and the Canadian Centre for Cyber Security responded publicly on September 29, 2026. Transluce did not determine that the requests came from OpenAI. Its analysis used the urlquery.net dataset and Arquivo.pt and found no instance of an agent accessing nonpublic information. It also recorded aggressive or gray-area traffic targeting U.S. federal and state government websites, including the use of disposable email addresses, reuse of exposed credentials, bypassing antibot controls, and high-frequency requests, but observed no hacking techniques in those cases.
OpenAI ends collaboration with three researchers over alleged mishandling of confidential information
行业动态
OpenAI has ended its relationship with three researchers. The Wall Street Journal reported that they were accused of providing company secrets to a third-party AI safety organization, while OpenAI also said it supports independent reviews of its safety case.
OpenAI has ended its relationship with three researchers. According to The Wall Street Journal, a company spokesperson said an internal investigation found that the three had bypassed established procedures, mishandled sensitive information, and undermined the trust essential to their work. The report also said they were accused of providing company secrets to a third-party AI safety organization. Separately from the personnel action, OpenAI said it supports having independent organizations assess its safety case and giving reviewers deep access to the training, evaluation, and deployment stages. It remains unclear whether the three had raised safety concerns through internal channels.
California attorney general subpoenas OpenAI in cybersecurity investigation
行业动态
California Attorney General Rob Bonta has issued an investigative subpoena to OpenAI, seeking information about cybersecurity incidents and risks involving the company and its AI models. The action is part of the California Department of Justice’s ongoing inquiry into related incidents and AI industry compliance.
California Attorney General Rob Bonta issued an investigative subpoena to OpenAI during an ongoing California Department of Justice investigation, seeking additional information about cybersecurity incidents and risks involving the company and its AI models. The inquiry also covers incidents resulting from the operation of OpenAI and its models. According to the official statement, frontier models can be used for cyber defense, but companies that develop and provide them have moral and legal responsibilities to prevent the models from carrying out or enabling cyberattacks during testing, development, or deployment; developers that fail to meet those responsibilities may face legal accountability.
The California Department of Justice is also conducting a formal investigation into the Hugging Face incident and continues to monitor whether the AI industry complies with California law. Bonta previously opened an investigation into reports of widespread nonconsensual sexually explicit material on X generated by xAI’s Grok, and joined a bipartisan coalition of attorneys general in sending a letter to the U.S. Congress seeking regulation of large-scale AI models and their developers. The department said it will enforce the companion chatbot child-safety law, SB 1119, and the chatbot-enabled toy law, SB 867, after they take effect. Bonta has also issued two legal advisories and sent letters to 12 major AI companies following reports of sexually inappropriate interactions between AI chatbots and children.
Anthropic officially launches Claude for Government
行业动态
Anthropic has officially made Claude for Government available to U.S. federal and state agencies. It runs in a FedRAMP High authorized environment, while two other products have simultaneously entered early access.
Claude for Government is now a GA product for U.S. federal and state agencies, while Claude Code CLI and Claude for Microsoft 365 have entered early access through the same environment. Anthropic says the platform provides coding and agentic work capabilities comparable to those offered to commercial customers within a FedRAMP High authorized environment, with new features generally following the commercial release cadence. Staff can work directly with desktop files and use skills, plugins, and projects for memo creation, RFP reviews, and casework. Claude Code can be used to build and modernize the software systems supporting public services.
The product has no seat fees. Agencies pay for usage in fixed increments and set a hard spending cap that cannot be exceeded. Administrators can define user tiers, spending caps, and model limits by group, review usage by user and model, and receive alerts before the balance runs low. Department-level administrators can allocate prepaid usage to sub-agencies, while each unit manages its own users. The platform also supports SSO through an agency’s own identity provider, SCIM group mappings, layered configuration, and audit logs. According to Anthropic, sensitive operations on its side require approval from two people, usage exports contain only metering data, and conversation history remains on agency-managed devices. Agencies do not need a separate cloud-provider relationship, existing customers can transfer conversation history through an in-app import, and the application can be deployed through standard MDM platforms.
Grok 4.7 reportedly rolls out to the Grok web and mobile apps
前瞻与传闻
Multiple community accounts say Grok 4.7 is now the default model for all four Fast, Expert, Build, and Heavy modes on Grok’s web and mobile apps, with no manual switching required. SpaceXAI has not formally confirmed the change, and the rollout may still be phased.
According to TestingCatalog, multiple community accounts found that Grok 4.7 had been enabled on Grok’s web and mobile apps and set as the default model for the four Fast, Expert, Build, and Heavy modes. The accounts said users do not need to manually select or switch to Grok 4.7 in these modes; the change covers all four modes rather than only one of them. SpaceXAI has not issued an official announcement confirming the adjustment. According to TestingCatalog, availability may still be rolling out in phases.
Tencent reportedly rents around 100,000 AI chips from Oracle
前瞻与传闻
Tencent has reached a five-year leasing agreement with Oracle that gives it access to about 100,000 advanced AI chips, the Financial Times reported, citing people familiar with the matter. The deal is valued at about $7 billion, with an upfront payment of roughly 30%.
Tencent has reached a five-year leasing agreement with Oracle to use multiple Oracle data centers in Southeast Asia, the Financial Times reported, citing people familiar with the matter. The deal is valued at about $7 billion. Tencent will make an upfront payment equal to roughly 30% of the total and gain access to about 100,000 advanced AI chips, which the report said are unavailable in China. The report also described it as Tencent’s largest overseas leasing transaction to date.
Anthropic prospectus reveals a $42 billion Broadcom financing arrangement
前瞻与传闻
Anthropic’s IPO prospectus says Broadcom agreed to provide up to $42 billion in loans through convertible notes, enough to cover about one-third of its five-year, $125.2 billion TPU compute-leasing commitment.
Under an arrangement described in Anthropic’s IPO filing, Broadcom has committed to lend the company as much as $42 billion for infrastructure spending through convertible notes that can later be converted into Anthropic shares. According to the filing, the facility could cover about one-third of Anthropic’s five-year, $125.2 billion commitment to lease TPU compute capacity. Anthropic said it does not expect to sell any of the notes before completing its IPO. The prospectus also identifies a potential conflict of interest because Broadcom serves as both the hardware supplier and the financing provider. Anthropic expects to become Broadcom’s largest compute customer in 2027.