Testing ads in ChatGPT
OpenAI announced a test of advertisements within ChatGPT, marking the company's exploration of ad-supported monetization models alongside its existing subscription offerings. The initiative aims to ba
Knowledge catalogue
OpenAI announced a test of advertisements within ChatGPT, marking the company's exploration of ad-supported monetization models alongside its existing subscription offerings. The initiative aims to ba
The Chrome extension expands what Codex can do for coding and work. From debugging browser flows to checking dashboards, conducting research, or updating CRMs, Codex can take on more of the tasks that
The most female-led product org in tech right now: Chief Product Officer: Ami Vora Claude Code/Cowork Head of Product: Cat Wu Claude Code/Cowork Head of Eng: Fiona Fung Claude Platform Head of Product
arXiv:2605.05029v1 Announce Type: new Abstract: We report a systematic failure mode in predictive representation learning. Across 2695 neural network configurations trained to predict linear-Gaussian
arXiv:2602.02315v2 Announce Type: replace Abstract: Large language models (LLMs) form implicit beliefs (posteriors over latent variables) from prompts, but we lack a mechanistic account of how these b
arXiv:2605.04481v1 Announce Type: new Abstract: Minimum-fuel low-thrust rendezvous guidance yields bang-bang control structures highly sensitive to estimation errors, sensor anomalies, and solver regu
arXiv:2605.04056v1 Announce Type: new Abstract: Representation learning seeks meaningful sensory representations without supervision and can model aspects of human development. Although many neural ne
arXiv:2605.04119v1 Announce Type: cross Abstract: Ancestral sequence reconstruction (ASR) aims to infer extinct protein sequences at internal nodes of a phylogenetic tree. Classical ASR methods are ty
arXiv:2605.04107v1 Announce Type: cross Abstract: Production agent frameworks (OpenAI Function Calling, Anthropic Tool Use, MCP) transmit tool schemas as JSON, a format designed for machine parsing, n
arXiv:2605.04409v1 Announce Type: new Abstract: Remote Sensing Image Change Captioning (RSICC) aims to generate spatially grounded natural language descriptions of scene evolution from bi-temporal ima
arXiv:2605.04941v1 Announce Type: new Abstract: This paper describes our system submitted to SemEval-2026 Task 11: Disentangling Content and Formal Reasoning in Large Language Models. We present an ef
arXiv:2605.05102v1 Announce Type: new Abstract: We study the distribution of regret in stochastic multi-armed bandits and episodic reinforcement learning through a unified framework. We formalize a di
U.S. intelligence says Iran can outlast Trump’s Hormuz blockade for months — via @washingtonpost https://www.washingtonpost.com/national-security/2026/05/07/cia-intelligence-iran-trump-blockade-missil
arXiv:2509.14448v2 Announce Type: replace Abstract: Benchmarks such as SWE-bench and ARC-AGI demonstrate how shared datasets accelerate progress toward artificial general intelligence (AGI). We introd
Vibe Inc., the creator of a contextual artificial intelligence workspace platform that handles real-world meetings, introduced a wearable device today that brings an AI assistant to professionals wher
arXiv:2506.06856v3 Announce Type: replace Abstract: Visual reasoning is crucial for understanding complex multimodal data and advancing Artificial General Intelligence. Existing methods enhance the re
arXiv:2605.04574v1 Announce Type: new Abstract: UAV-ground visual tracking (UGVT) aims to simultaneously track the same object from both the UAV and the ground view. However, existing two-stream metho
arXiv:2605.04870v1 Announce Type: new Abstract: Video text-based visual question answering (Video TextVQA) aims to answer questions by reasoning over visual textual content appearing in videos. Despit
arXiv:2605.05161v1 Announce Type: new Abstract: Zero-shot anomaly localisation via vision-language models (VLMs) offers a compelling approach for rare pathology detection, yet its performance is funda
We already had gemini-3.1-flash-lite-preview back on March 3rd, not clear if this new gemini-3.1-flash-lite is different other than no longer being marked as a 'preview'. Pricing appears to be the sam
OpenAI announced on X that voice feature updates for ChatGPT are in development and coming soon, though no specific timeline was provided. The post acknowledges user demand for enhanced voice capabili
We recently found some instances of CoT grading during the training of previously deployed models after building a system that scans all OpenAI RL runs for accidental CoT grading. We did not find clea
Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash. https://github.com/antirez/ds4 This project would have been impossible without the existence of llama.cpp and GGML and the work of
We've teamed up with @cerebras to offer free Windsurf plans for SWE-1.6 Fast Mode at up to 1000 tok/s! Fast Mode is built on Cerebras inference, enabling superior speed for planning and development wi
Boris Cherny praised an Anthropic AI event that brought together a doctors' coding community, expressing enthusiasm about the gathering's energy and engagement. The post suggests Anthropic hosted or s
arXiv:2605.03354v1 Announce Type: new Abstract: Agent memory failures are silent: an LLM-based agent can produce a fluent response even when it fails to extract, retain, or retrieve the information ne
Simon Willison / Simon Willison's Weblog: While Anthropic will use the Colossus 1 data center, which has a really bad environmental record, xAI retains the larger Colossus 2 for its own AI training —
who’s adding this to reachy mini? Introducing GPT-Realtime-2 in the API: our most intelligent voice model yet, bringing GPT-5-class reasoning to voice agents. Voice agents are now real-time collaborat
The Firefox development team reportedly resolved a significantly higher volume of security vulnerabilities in April with assistance from Claude Mythos Preview, an AI tool, compared to their bug-fixing
This post shows a Wordle game result where the player solved puzzle #1,782 in 3 attempts, displaying the emoji grid that represents their guesses and letter placements (gray for incorrect letters, yel
This post documents a Wordle game completion (puzzle #1,783) where the player solved the word in six attempts, the maximum allowed before failure. The color-coded emoji grid shows the progression of g
arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able
arXiv:2605.01546v1 Announce Type: cross Abstract: Sixth-generation (6G) networks are increasingly envisioned as AI-native infrastructures integrating communication, sensing, and computing into a unifi
arXiv:2605.03941v1 Announce Type: new Abstract: Achieving Artificial General Intelligence (AGI) requires agents that learn and interact adaptively, with interactive world models providing scalable env
A critical question in agent design is “how do we build agentic workflows so humans are given significant, interesting, or variance-producing decisions as they come up in the work?” A Claude-run compa
arXiv:2605.03857v1 Announce Type: new Abstract: This work presents a deeper analysis of the 'irreversibility' property of PolyProtect, a biometric template protection method initially proposed for sec
arXiv:2605.03832v1 Announce Type: new Abstract: In recent years, machine learning has made significant progress in clinical outcome prediction, demonstrating increasingly accurate results. However, th
arXiv:2605.02942v1 Announce Type: cross Abstract: Bias in medical AI is often framed as a problem of representation. However, in image-based tasks such as fetal ultrasound, performance disparities can
arXiv:2602.18843v3 Announce Type: replace Abstract: We introduce ABD, a benchmark for default-exception abduction over finite first-order worlds. Given a background theory with an abnormality predicat
arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw
arXiv:2605.03076v1 Announce Type: new Abstract: Graph contrastive learning (GCL) has become a central paradigm for self-supervised representation learning in computational intelligence, with applicati
arXiv:2601.06395v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly multilingual, yet open models continue to underperform relative to proprietary systems, with the gap m
arXiv:2605.03590v1 Announce Type: new Abstract: Recent large language models (LLMs) show strong speech recognition and translation capabilities for high-resource languages. However, African languages
arXiv:2605.03808v1 Announce Type: cross Abstract: Agentic data science (ADS) systems are rapidly improving their capability to autonomously analyze, fit, and interpret data, potentially moving towards
arXiv:2602.12631v2 Announce Type: replace-cross Abstract: Inventory control is a fundamental operations problem in which ordering decisions are traditionally guided by theoretically grounded operation
arXiv:2605.02738v1 Announce Type: new Abstract: Solar photovoltaic (PV) deployment is expanding rapidly, yet detailed, up-to-date information on the spatial distribution and capacity of rooftop PV rem
arXiv:2605.01415v1 Announce Type: new Abstract: Recent AI systems compress the distance between capability growth and capability deployment. Earlier high-risk technologies were slowed by capital inten
arXiv:2605.03710v1 Announce Type: cross Abstract: Bayesian predictive inference propagates parameter uncertainty to quantities of interest through the posterior-predictive distribution. In practice, t
arXiv:2605.02669v1 Announce Type: new Abstract: Drug-induced liver injury (DILI) remains a leading cause of late-stage clinical trial attrition. However, existing computational predictors primarily re
arXiv:2605.03624v1 Announce Type: new Abstract: Aspect-Based Sentiment Analysis (ABSA) enables fine-grained opinion analysis by identifying sentiments toward specific aspects or targets within a text.
Anthropic has signed an agreement with SpaceX to use all of the compute capacity at SpaceX's Colossus 1 data center in Memphis, Tennessee, providing access to over 300 megawatts of capacity (220,000+
Anthropic PBC today announced that it will use SpaceX Corp.’s Colossus 1 supercomputer to power its Claude chatbot. The system was originally built in 2024 by xAI Holdings Corp., an artificial intelli
arXiv:2605.01727v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for automated news credibility assessment, yet it remains unclear whether they apply even-handed stan
arXiv:2605.01420v1 Announce Type: new Abstract: Artificial Jagged Intelligence (AJI) denotes a recurring pattern in which large learning systems exhibit strong local capabilities while remaining weak
arXiv:2605.02967v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances LLMs, but performance is highly sensitive to complex architecture designs and hyper-parameter configurat
arXiv:2605.03496v1 Announce Type: new Abstract: We consider function optimization as a sequential decision making problem under budget constraint. This constraint limits the number of objective functi
arXiv:2605.03759v1 Announce Type: new Abstract: While Large Vision-Language Models (LVLMs) offer powerful capabilities, they pose privacy risks by unintentionally memorizing sensitive personal informa
arXiv:2512.09874v2 Announce Type: replace Abstract: Correctly parsing mathematical formulas from PDFs is critical for training large language models and building scientific knowledge bases from academ
arXiv:2605.03742v1 Announce Type: new Abstract: This paper is devoted to the adaptation of generative large language models for the Tajik language, a low-resource language with Cyrillic script. To ove
arXiv:2605.03509v1 Announce Type: new Abstract: Low-light image enhancement is a fundamental challenge in computer vision and multimedia applications, as images captured under insufficient illuminatio