AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlog
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,060 results
Safety

When and How Long? The Readout-Mediator Angle in Temporal Reasoning

DGX agent

arXiv:2605.29126v1 Announce Type: cross Abstract: A linear probe can decode a representation almost perfectly and yet be completely irrelevant to how the model uses it. On calendar-date duration reaso

safetyarxiv-cs-ai
29 May 2026
Local Ai

When Cloud Agents Meet Device Agents: Lessons from Hybrid Multi-Agent Systems

X Post
Paper
YouTube
Reddit
GitHub
DGX agent

arXiv:2605.30102v1 Announce Type: cross Abstract: The design space of agentic AI inference spans two extremes: frontier large language models (LLMs), typically hosted in the cloud and offering strong

local-aiarxiv-cs-ai
29 May 2026
Research

When Do Graph Foundation Models Transfer? A Data-Centric Theory

DGX agent

arXiv:2605.29828v1 Announce Type: new Abstract: Graph foundation models (GFMs) aim to reuse a single backbone across diverse graph domains, yet their transfer is often uneven and can exhibit negative

researcharxiv-cs-lg
29 May 2026
Applications

When Does Persona Prompting Actually Help? A Retrieval and Metric Analysis of Expert Role Injection in LLMs

DGX agent

arXiv:2605.29420v1 Announce Type: new Abstract: Persona prompting is widely used to steer large language models, yet its practical value remains unclear. Prior work often evaluates persona prompting u

applicationsarxiv-cs-ai
29 May 2026
Industry

'When I started SpaceX, one of my friends got a compilation of rocket failures & made me watch the whole thing. I knew the probability of Sp…

DGX agent

Elon Musk recounts how a friend showed him a compilation of rocket failures when he was starting SpaceX, making him aware of the inherent risks and high failure probability in the aerospace industry.

industryelon-musk--x
29 May 2026
Model Releases

When LLM Reward Design Fails: Diagnostic-Driven Refinement for Sparse Structured RL

DGX agent

arXiv:2605.28918v1 Announce Type: new Abstract: For sparse, structured reinforcement-learning tasks with semantic reward-function interfaces, LLM-generated reward shaping is better framed as debugging

model-releasesarxiv-cs-lg
29 May 2026
Research

When Models Disagree: Rethinking LLM Evaluation for Public Comment Analysis

DGX agent

arXiv:2605.29025v1 Announce Type: new Abstract: Federal agencies are deploying large language models (LLMs) to categorize public comment corpora, where the model's organization of the record shapes wh

researcharxiv-cs-ai
29 May 2026
Tutorials

When Models Learn to Ask Why: Adaptive Causal Reasoning for Trustworthy Medical Vision-Language Models

DGX agent

arXiv:2603.23085v2 Announce Type: replace Abstract: Vision-Language Models (VLMs) have enabled interpretable medical diagnosis by integrating visual perception with linguistic reasoning. Yet, existing

tutorialsarxiv-cs-ai
29 May 2026
Research

When RL Suppresses Its Own Vocabulary: Recovering Reasoning Diversity in Puzzle-to-Math Transfer

DGX agent

arXiv:2605.29190v1 Announce Type: cross Abstract: Reinforcement learning using verifiable rewards (RLVR) improves LLM reasoning, but the conditions under which it transfers across domains -- and why i

researcharxiv-cs-cl
29 May 2026
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When Should Models Change Their Minds? Contextual Belief Management in Large Language Models

DGX agent

arXiv:2605.30219v1 Announce Type: new Abstract: Long-horizon interactions require language models to manage accumulating information: when to update their state, when to preserve their state, and what

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When the Same Coefficients Reach Different Places: Asymmetric Realizability in Transplanting Tokenizers across Large Language Models

DGX agent

arXiv:2601.00065v3 Announce Type: replace-cross Abstract: Tokenizer transplant in cross-vocabulary model composition reconstructs donor-only embedding rows as weighted combinations over shared lexical

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, whic…

DGX agent

When we say “LiteParse runs everywhere,” we mean it. Our WASM package is lightweight, minimal, and built for browser and edge runtimes, which makes it a perfect fit for @cloudflare Workers. Using WebA

model-releasesjerry-liu--x
29 May 2026
Research

When, why, and how do diffusion posterior samplers fail? A finite-sample lens

DGX agent

arXiv:2605.30330v1 Announce Type: new Abstract: Diffusion models have excellent capacity to model complex distributions of natural data, which has made them a popular and effective choice for posterio

researcharxiv-cs-lg
29 May 2026
Model Releases

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being.

DGX agent

When you are talking to an LLM, you are speaking to a synthesized work of interactive fiction, not a real being. ChatGPT, Claude, and Sydney are not their neural networks. If any LLM claims to be cons

model-releasesgary-marcus--x
29 May 2026
Applications

Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues

DGX agent

arXiv:2605.30051v1 Announce Type: new Abstract: A key part of developing large language model (LLM)-powered, automated tutoring tools is student simulation, i.e., using LLMs to role-play as students,

applicationsarxiv-cs-cl
29 May 2026
Model Releases

Who can we trust? LLM-as-a-jury for Comparative Assessment

DGX agent

arXiv:2602.16610v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly applied as automatic evaluators for natural language generation assessment often using pairwise

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Why Far Looks Up: Probing Spatial Representation in Vision-Language Models

DGX agent

arXiv:2605.30161v1 Announce Type: new Abstract: Vision-language models (VLMs) achieve strong performance on spatial reasoning benchmarks, yet it remains unclear whether this reflects structured 3D und

model-releasesarxiv-cs-cv
29 May 2026
Tutorials

Why Larger Models Learn More: Effects of Capacity, Interference, and Rare-Task Retention

DGX agent

arXiv:2605.29548v1 Announce Type: new Abstract: Larger models learn tasks smaller models do not. What drives this phenomenon? We develop a simple phenomenological argument that power-law scaling alrea

tutorialsarxiv-cs-lg
29 May 2026
Model Releases

Why Specialist Models Still Matter: A Heterogeneous Multi-Agent Paradigm for Medical Artificial Intelligence

DGX agent

arXiv:2605.29744v1 Announce Type: new Abstract: The impressive performance of generalist large language models (LLMs) such as GPT and Claude in healthcare raises a critical question: will domain-speci

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows su…

DGX agent

Windows users, this one’s for you. Computer use now works on Windows, so Codex can take action on your Windows computer. And with Windows support for Codex in the ChatGPT mobile app, you can start, re

model-releasesopenai--x
29 May 2026
Industry

Winning under CMS TEAM: Building the learning health system to realize success in VBC today and tomorrow

DGX agent

This resource discusses how healthcare organizations can build learning health systems to succeed under CMS TEAM (Transforming Episode Accountability Models) and advance value-based care (VBC) initiat

industrydatabricks
29 May 2026
Research

Woah. A profound shift in American science is coming. Every federal research grant could soon require sign-off from a political appointee. S…

DGX agent

Woah. A profound shift in American science is coming. Every federal research grant could soon require sign-off from a political appointee. Scientific progress depends on funding decisions grounded in

researchyann-lecun--x
29 May 2026
Model Releases

Wordle 1,804 4/6 ⬛🟨⬛⬛⬛ ⬛🟩⬛⬛🟨 ⬛🟩🟩🟩🟩 🟩🟩🟩🟩🟩

DGX agent

This entry documents a Wordle game result where Anthropic solved puzzle #1,804 in 4 attempts, using color-coded feedback (⬛ = incorrect letter, 🟨 = correct letter wrong position, 🟩 = correct letter co

model-releasesanthropic--x
29 May 2026
Agents

Worked on some code this morning using Opus 4.8 and so far I'm really liking it. Much more cooperative than 4.7 and less 'over agentic'. Sto…

DGX agent

Worked on some code this morning using Opus 4.8 and so far I'm really liking it. Much more cooperative than 4.7 and less 'over agentic'. Stops and asks for my input when needed in places 4.7 (and GPT

agentsjeremy-howard--x
29 May 2026
Tools

Working at the intersection of AI, data, and the built environment, few have shaped the language of AI-driven art like @refikanadol. See him…

DGX agent

Working at the intersection of AI, data, and the built environment, few have shaped the language of AI-driven art like @refikanadol. See him take the stage with @pirroh on day two of Vibecon. NYC, Jun

toolsreplit--x
29 May 2026
Model Releases

World Models in Words: Auditing Physical State-Transition Commitments in Vision-Language Models

DGX agent

arXiv:2605.29585v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly used to answer questions about physical scenes, yet most evaluations reduce performance to a final answer

model-releasesarxiv-cs-cl
29 May 2026
Agents

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

DGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

agentsarxiv-cs-cl
29 May 2026
Research

X-GS: An Extensible Framework for Perceiving and Thinking via 3D Gaussian Splatting

DGX agent

arXiv:2603.09632v3 Announce Type: replace-cross Abstract: 3D Gaussian Splatting (3DGS) has emerged as a powerful technique for novel view synthesis, subsequently extending into numerous spatial AI app

researcharxiv-cs-cl
29 May 2026
Industry

XCENA raises $135M for its computational memory controller

DGX agent

XCENA Inc., a startup with a memory device designed to speed up artificial intelligence clusters, today announced that it has raised 135 million in funding. The Series B round was led by Korean funds

industrysiliconangle
29 May 2026
Industry

Xcena, whose MX1 chip performs data orchestration and KV cache management directly within memory modules, raised a 135M Series B at a 570M valuation (Kate Park/TechCrunch)

DGX agent

Kate Park / TechCrunch: Xcena, whose MX1 chip performs data orchestration and KV cache management directly within memory modules, raised a 135M Series B at a 570M valuation — Every time you ask ChatGP

industrytechmeme
29 May 2026
Research

Xetrieval: Mechanistically Explaining Dense Retrieval

DGX agent

arXiv:2605.29507v1 Announce Type: new Abstract: Explaining why dense retrievers assign high relevance scores remains challenging because retrieval decisions are made through opaque high-dimensional em

researcharxiv-cs-ai
29 May 2026
Safety

xModel-KD: Cross-modal Knowledge Distillation for 3D Scene Perception using LiDAR

DGX agent

arXiv:2605.30111v1 Announce Type: cross Abstract: Point cloud segmentation is a fundamental task in 3D scene understanding. Its progress is constrained by the high cost and time required for dense 3D

safetyarxiv-cs-ai
29 May 2026
Safety

yes, absolutely, many companies are experimenting. but also: most of those experiments are failing to yield significant RoI. (weird for an e…

DGX agent

yes, absolutely, many companies are experimenting. but also: most of those experiments are failing to yield significant RoI. (weird for an economist to not even ask or address that question.) Looks li

safetygary-marcus--x
29 May 2026
Research

.@ylecun’s definition of what is a world model.

DGX agent

Yann LeCun, a pioneering AI researcher, provides his definition of what constitutes a world model in this X post. A world model is an AI system's internal representation of how the physical world work

researchyann-lecun--x
29 May 2026
Model Releases

YoCausal: How Far is Video Generation from World Model? A Causality Perspective

DGX agent

arXiv:2605.30346v1 Announce Type: new Abstract: As video diffusion models (VDMs) advance toward world models, a key question arises: do they truly understand causality, or merely overfit to statistica

model-releasesarxiv-cs-cv
29 May 2026
Safety

you break it, you buy it peter thiel has broken the united states, and now he is abandoning it

DGX agent

you break it, you buy it peter thiel has broken the united states, and now he is abandoning it Peter Thiel has temporarily relocated his family to Argentina, enrolled his children in school there, and

safetygary-marcus--x
29 May 2026
Industry

You pick AI models based on their model cards, why not humans haha. Well done Noah! https://huggingface.co/noahmclaughlin/Noah-McLaughlin-7B

DGX agent

Noah McLaughlin created a 7-billion parameter language model and published it on Hugging Face, drawing a humorous parallel to how AI practitioners evaluate models using model cards by suggesting human

industryclem-delangue--x
29 May 2026
Tutorials

Zero-shot CT Super-Resolution using Diffusion-based 2D Projection Priors and Signed 3D Gaussians

DGX agent

arXiv:2508.15151v3 Announce Type: replace-cross Abstract: Computed tomography (CT) is important in clinical diagnosis, but acquiring high-resolution (HR) CT is constrained by radiation exposure risks.

tutorialsarxiv-cs-cv
29 May 2026
Safety

1. Agreed that OpenAI is in deep trouble; that’s why I have long suggested that it might be the WeWork of AI but 2. Anthropic is not out of …

DGX agent

1. Agreed that OpenAI is in deep trouble; that’s why I have long suggested that it might be the WeWork of AI but 2. Anthropic is not out of the woods; their best quarter was exactly when tokenmaxxing

safetygary-marcus--x
28 May 2026
Industry

2027 Audi RS5 first drive: A performance PHEV with split personalities

DGX agent

The 2027 Audi RS5 is Audi Sport's first high-performance plug-in hybrid, pairing maximum performance with remarkable efficiency. The powertrain combines a 2.9-liter twin-turbo V6 with a 130-kW electri

industryars-technica
28 May 2026
Agents

3 hour turn around on custom feature request!

DGX agent

3 hour turn around on custom feature request! @yoheinakajima @cura_inc okay live! custom tags / labels so you can filter as you see fit! configurable on the UI but also via MCP/AI so super easy for yo

agentsyohei-nakajima--x
28 May 2026
Tools

3. Keep Secrets Server-Side API keys, tokens, and database URLs in client-side code, localStorage, or cookies are basically public. Anyone c…

DGX agent

3. Keep Secrets Server-Side API keys, tokens, and database URLs in client-side code, localStorage, or cookies are basically public. Anyone can open dev tools and grab them. Use Replit Secrets to store

toolsreplit--x
28 May 2026
Tools

4. Secure Your Users Rolling your own auth means a dozen ways to leak data, from weak password hashing to broken session handling to missing…

DGX agent

4. Secure Your Users Rolling your own auth means a dozen ways to leak data, from weak password hashing to broken session handling to missing rate limits. Use Replit Auth or Clerk instead. They handle

toolsreplit--x
28 May 2026
Model Releases

$65B private round More than double the size of the largest IPO ever

DGX agent

65B private round More than double the size of the largest IPO ever We've raised 65 billion in Series H funding at a $965 billion post-money valuation, led by @AltimeterCap, Dragoneer, @Greenoaks, and

model-releasesjeremy-howard--x
28 May 2026
Industry

A $2,000 AI-generated film will make its debut at Tribeca

DGX agent

Next month's Tribeca Festival will include the premiere of an AI-generated film: Dreams of Violets. The 75-minute film is a fictional dramatization of the Iranian government's mass killing of protesto

industrythe-verge-ai
28 May 2026
Model Releases

A Bayesian Nonparametric Perspective on Mahalanobis Distance for Out of Distribution Detection

DGX agent

arXiv:2502.08695v2 Announce Type: replace-cross Abstract: Bayesian nonparametric methods are naturally suited to the problem of out-of-distribution (OOD) detection. However, these techniques have larg

model-releasesarxiv-cs-lg
28 May 2026
Model Releases

A Broader View of Thompson Sampling

DGX agent

arXiv:2510.07208v2 Announce Type: replace Abstract: Thompson Sampling is one of the most widely used and studied bandit algorithms, known for its simple structure, low regret performance, and solid th

model-releasesarxiv-cs-lg
28 May 2026
← Previous
1…10101011101210131014…1898
Next →