AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,914
  • Agents7,747
  • Applications5,533
  • Concepts5
  • Hardware1,920
  • Industry6,196
  • Local Ai5,094
  • Model Releases24,727
  • Research20,781
  • Safety13,739
  • Syntheses17
  • Tools1,678
  • Tutorials3,477

Source
Human
90,914Total entries
1Added by human
90,913Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,913 results
2 Jun 2026

The Image Reconstruction Game: Drawing Common Ground Through Iterative Multimodal Dialogue

Model ReleasesDGX agent

arXiv:2606.01901v1 Announce Type: cross Abstract: We introduce the Image Reconstruction Game, a fully automated benchmark in which a vision-language model issues corrective instructions to an image ge

The Invisible Coalition Partner: How LLMs Vote When Democracy Gets Concrete

Model ReleasesDGX agent

arXiv:2606.00048v1 Announce Type: cross Abstract: Prior research has established that instruction-tuned large language models exhibit left-of-center political bias, measured exclusively through abstra

The IPOs of SpaceX, Anthropic, and OpenAI could add up to $4T to US stock market value within months, fueling concerns they could trigger more capital-raising (The Economist)

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Applications
DGX agent

The Economist: The IPOs of SpaceX, Anthropic, and OpenAI could add up to 4T to US stock market value within months, fueling concerns they could trigger more capital-raising — They promise to be the bi

The Lie We Tell: Correcting the Euclidean Fallacy in Vision Language Action Policies via Score Matching on Tangent Space

ResearchDGX agent

arXiv:2606.01847v1 Announce Type: cross Abstract: Diffusion-based Vision-Language-Action policies achieve remarkable success in robotic manipulation, yet commit a fundamental geometric error we term t

The New Social Image: How AI Competency and AI Proactivity Influence Self- and Peer-Perceptions in the Workplace

ResearchDGX agent

arXiv:2606.00182v1 Announce Type: cross Abstract: Human-AI collaboration is considered the most promising way to incorporate AI in the workplace. What remains unexplored are the experiential consequen

The next evolution of Hermes Agent is here! Introducing Hermes Desktop: everything you love about Hermes, now native on your machine. First …

AgentsDGX agent

The next evolution of Hermes Agent is here! Introducing Hermes Desktop: everything you love about Hermes, now native on your machine. First demoed in Jensen's GTC keynote, it's now in public preview.

The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs

SafetyDGX agent

arXiv:2606.00674v1 Announce Type: cross Abstract: Large Language Models (LLMs) aligned via outcome-based Reinforcement Learning (RL) frequently exhibit a critical failure mode: they achieve high perfo

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth takin…

SafetyDGX agent

The persistent and possibly industry-funded campaign to discredit me begins with a fundamental misunderstanding of my work. It’s worth taking the time to understand the issues, if you want to understa

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the…

SafetyDGX agent

The persistent campaign to discredit me begins with a fundamental misunderstanding of my work. It is worth taking the time to understand the issues, if you want to understand a lot of what is driving

“The public has swung 49 points against data centers in just nine months, underscoring the heightened political salience of the facilities a…

SafetyDGX agent

Public opinion has shifted dramatically against data centers over a nine-month period, with a 49-point swing in unfavorable sentiment, reflecting growing political concern about these facilities. This

The quarq agent is built on LangGraph! LangGraph makes it easy to build complex memory systems (quarq is now at the top of the LongMemEval l…

AgentsDGX agent

The Quarq agent is constructed using LangGraph, a framework that simplifies the development of complex memory systems. Quarq has achieved a top ranking on the LongMemEval benchmark, demonstrating the

The Refusal--Compliance Tradeoff: A Large-Scale Safety Behavior Audit of Large Language Models

Model ReleasesDGX agent

arXiv:2605.05427v2 Announce Type: replace Abstract: Refusal rates are a poor proxy for LLM safety, i.e., a model may over-refuse benign prompts while still complying with harmful ones. We audit both f

The Representation-Rationalizability Tradeoff in Reward Learning

ResearchDGX agent

arXiv:2606.00291v1 Announce Type: cross Abstract: In RLHF, each training example contains a prompt x and two candidate responses y,y', and annotators provide pairwise preferences between these respons

The Right Inference Strategy Is All You Need: Nearly Training-Free Domain-Wise Inference for EgoCross Challenge

ResearchDGX agent

arXiv:2606.00829v1 Announce Type: new Abstract: EgoCross evaluates multimodal large language models on egocentric video question answering under substantial domain shift, where test videos come from s

The Role of Ambiguity in Error Prediction via Uncertainty Quantification

ResearchDGX agent

arXiv:2606.02093v1 Announce Type: cross Abstract: The task of Error Prediction, namely predicting whether a model output is correct, is commonly tackled with Uncertainty Quantification (UQ). However,

The role of class encoding in neural collapse

SafetyDGX agent

arXiv:2606.00344v1 Announce Type: new Abstract: Neural collapse is a structural property of the last-hidden-layer activations in neural network classification models, when trained beyond a zero classi

The Shape of Wisdom: Decision Trajectories in Language Models

Model ReleasesDGX agent

arXiv:2606.01202v1 Announce Type: new Abstract: Language models do not simply choose an answer at the output layer. In a 9,000-trajectory MMLU study across Qwen2.5-7B-Instruct, Llama-3.1-8B-Instruct,

The Social Cost of Intelligence: Emergence, Propagation, and Amplification of Stereotypical Bias in Multi-Agent Systems

SafetyDGX agent

arXiv:2510.10943v2 Announce Type: replace-cross Abstract: Bias in large language models (LLMs) remains a persistent challenge, often leading to stereotyping and unfair treatment across social groups.

The US sanctions Nobitex, Iran's largest crypto exchange, accusing it of helping Iran's government and blacklisted state institutions evade Western sanctions (Gavin Finch/Reuters)

IndustryDGX agent

Gavin Finch / Reuters: The US sanctions Nobitex, Iran's largest crypto exchange, accusing it of helping Iran's government and blacklisted state institutions evade Western sanctions — The United States

Theoretical Analysis of Engression and Reverse Markov Engression

ResearchDGX agent

arXiv:2606.01002v1 Announce Type: cross Abstract: Engression is a recently proposed and effective framework for conditional distribution learning. Its multi-step Reverse Markov extension further impro

There's never been anything like the Henry Nowak case where the victim is non-white and the assailant is white, and the police respond the w…

IndustryDGX agent

There's never been anything like the Henry Nowak case where the victim is non-white and the assailant is white, and the police respond the way they did. That whole combination of events could really o

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to!

SafetyDGX agent

Things are so chaotic in AI right now that @axios had to specify *which* AI backlash they were referring to! Anthropic faces AI spending backlash before IPO https://www.axios.com/2026/06/02/anthropic-

Thinking Economically: A Hierarchical Framework for Adaptive-Complexity Reasoning in LLMs

ResearchDGX agent

arXiv:2606.01168v1 Announce Type: new Abstract: Chain-of-Thought (CoT) has significantly enhanced LLM reasoning, yet often incurs substantial computational overhead due to 'overthinking': generating e

Thinking in Blender: Staged Executable Inverse Graphics with Vision-Language Models

AgentsDGX agent

arXiv:2606.02580v1 Announce Type: new Abstract: Inverse graphics is a longstanding and highly underconstrained problem that seeks to reconstruct images as editable 3D scenes which can be rendered, rel

ThinkSwitch: Context Distillation with LoRA and Weight Interpolation for Specific-Purpose Reasoning Tasks

ResearchDGX agent

arXiv:2606.01080v1 Announce Type: cross Abstract: Large language models often improve on difficult tasks by spending inference-time compute on a reasoning trace before producing the final answer. That

This is actually one of the main advantages startups have over frontier labs, as long as there's a healthy spectrum of open-weight to closed…

Model ReleasesDGX agent

This is actually one of the main advantages startups have over frontier labs, as long as there's a healthy spectrum of open-weight to closed-weight models on the cost-performance curve. Building a mod

This is also available on the Claude Blog! https://claude.com/blog/a-harness-for-every-task-dynamic-workflows-in-claude-code

Model ReleasesDGX agent

This post discusses dynamic workflows in Claude Code, enabling flexible task automation and execution patterns. The content is available on the official Claude Blog and addresses how Claude can be har

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use age…

Model ReleasesDGX agent

🌞This is big Local AI news! A new open-source Computer-Use LLM has just launched. Holo 3.1 is H Company’s (🇫🇷) new local computer-use agent model that beats Qwen3.5-397B, Kimi-K2.5, and Sonnet 4.6! Si

This is disaster from Meta AI. Imagine being able to hack high profile accounts like White House, the U.S. Space Force, and Sephora simply b…

IndustryDGX agent

This is disaster from Meta AI. Imagine being able to hack high profile accounts like White House, the U.S. Space Force, and Sephora simply by chatting with a support bot. why would an AI chatbot be al

this is fine 🐶☕️🔥

Model ReleasesDGX agent

This post likely references the popular 'This is Fine' meme featuring a dog in a burning room, often used to comment on problematic situations being accepted or ignored. Without access to the specific

This Thursday: Idea → Money. @raymmar and I are taking a real project live, all the way from prototype to monetized: 1. Prototype the idea i…

ToolsDGX agent

This Thursday: Idea → Money. @raymmar and I are taking a real project live, all the way from prototype to monetized: 1. Prototype the idea into a working build 2. Build in the monetization 3. Set up t

THRD: A Training-Free Multi-Turn Defense Framework for Jailbreak Attacks on Large Language Models

SafetyDGX agent

arXiv:2606.01738v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks pose a growing threat to LLMs by exploiting conversational dynamics such as gradual escalation and cross-turn coordinatio

Threading Optimization for Vision-Language-Action Model Inference in Low-Cost Smart Agricultural Manipulation

SafetyDGX agent

arXiv:2606.00966v1 Announce Type: new Abstract: Vision-Language Action (VLA) models continue to face challenges such as slow inference speed and difficulty performing fine-grained motion adjustments,

Threshold-Based Exclusive Batching for LLM Inference

HardwareDGX agent

arXiv:2606.00516v1 Announce Type: new Abstract: Mixed batching (MB)--interleaving prefill and decode in a single batch--has become the standard scheduling strategy for large language model (LLM) infer

Thrive Holdings, a spinoff of Thrive Capital, commits $1B to acquire local accounting firms through its subsidiary, Current, and use AI to automate them (Anna Tong/Forbes)

Local AiDGX agent

Anna Tong / Forbes: Thrive Holdings, a spinoff of Thrive Capital, commits $1B to acquire local accounting firms through its subsidiary, Current, and use AI to automate them — In Thrive Holdings' live-

Through the PRISM: Principle-Aware, Interpretable, and Multi-Scale Evaluation of Visual Designs

Model ReleasesDGX agent

arXiv:2606.00592v1 Announce Type: new Abstract: Effective visual communication stems from the harmony of multiple design principles, such as readability, contrast, alignment, overlap, and coherence, w

TIDES: Time-Derivative Event Simulation via Deformable Reconstruction

TutorialsDGX agent

arXiv:2606.02058v1 Announce Type: new Abstract: Event cameras emit asynchronous events in response to environmental appearance changes. The scarcity of real-world event datasets makes simulation essen

TIGER: Traceable Inference with Graph-Based Evidence Routing for Mitigating Hallucinations in Multimodal Generation

Local AiDGX agent

arXiv:2606.00232v1 Announce Type: new Abstract: We study fact-level repair for multimodal generation, where a fluent output may contain specific facts that are not supported by the input. Existing inf

Time-Aware Diffusion based on Preference Disentanglement for Generative Recommendation

ApplicationsDGX agent

arXiv:2606.01670v1 Announce Type: cross Abstract: Recently, Generative Recommenders (GRs) have emerged as a transformative recommendation paradigm by replacing traditional item IDs with semantic indic

Time-Optimal Collision Avoidance Via a Greedy Polynomial Backward Sweep

Model ReleasesDGX agent

arXiv:2606.01169v1 Announce Type: cross Abstract: Spacecraft collision avoidance for low-thrust satellites often requires determining not only how to maneuver, but also how late a maneuver can begin w

TimeBlocks: Foundational and Continual Time-Series Blockbase -- Extended Version

ResearchDGX agent

arXiv:2606.02142v1 Announce Type: new Abstract: The ongoing digitization has led to a proliferation of time-series data streams that monitor a variety of processes, from which valuable insights may be

TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning

Model ReleasesDGX agent

arXiv:2606.01498v1 Announce Type: cross Abstract: Time series data inform critical decisions across many real-world domains. While large language model (LLM) agents can analyze data through natural la

Tiny Recursive Models for Solving the J2-Perturbed Lambert Problem

Model ReleasesDGX agent

arXiv:2606.00895v1 Announce Type: cross Abstract: This paper presents a fast, recursive neural solver for the J2-perturbed Lambert problem based on Tiny Recursive Models (TRM), termed the TRM-Perturbe

title undersells it - this @workos talk is doing v well and is the first to seriously challenge @mattpocockuk in weeks. team is ab testing

TutorialsDGX agent

title undersells it - this @workos talk is doing v well and is the first to seriously challenge @mattpocockuk in weeks. team is ab testing My talk from AIE Europe is up! Come learn the lessons I learn

TLG: Temporal-Logic Grounding for Video Question Answering via Source-Annotation Reconstruction and Category-Targeted Reasoning

Model ReleasesDGX agent

arXiv:2606.01591v1 Announce Type: new Abstract: The TimeLogic Challenge evaluates formal temporal-logic reasoning over video - 16 operators (before, after, until, since, always, co-occur, ordering, ..

TN-SHAP-G: Graph-Structured Tensor Network Surrogates for Shapley Values and Interactions

ResearchDGX agent

arXiv:2606.01540v1 Announce Type: cross Abstract: Shapley values are a widely used tool for attributing importance and interactions among input variables in black-box models, but their computation inv

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most i…

SafetyDGX agent

🚨 Today is a milestone in US AI policy, and an unbelievable moment for me, personally. 🚨 Here’s what I told @senjohnkennedy was the most important policy to implement, at the US Senate in May 2023, an

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model runnin…

Local AiDGX agent

Today we're announcing that hybrid agentic inference is coming to Perplexity Computer. Computer can split tasks between a local model running on your machine and frontier models in the cloud. This kee

Today’s question. Why hasn’t Vickrum Digwa’s brother (Gurpreet Digwa) been arrested and charged with assisting an offender? This lying POS c…

IndustryDGX agent

Today’s question. Why hasn’t Vickrum Digwa’s brother (Gurpreet Digwa) been arrested and charged with assisting an offender? This lying POS called 999 and said that Henry attacked his brother and racia

Token Predictors Are Not Planners: Building Physically Grounded Causal Reasoners

Model ReleasesDGX agent

arXiv:2606.01810v1 Announce Type: new Abstract: Current benchmarks for embodied vision-language planning often favor linguistic next-token prediction over physically grounded next-state reasoning. Thi

ToMAP: Training Opponent-Aware LLM Persuaders with Theory of Mind

TutorialsDGX agent

arXiv:2505.22961v3 Announce Type: replace Abstract: Large language models (LLMs) have shown promising potential in persuasion, but existing works on training LLM persuaders are still preliminary. Nota

ToolFG: Towards Well-Grounded Fine-Grained Image Classification

SafetyDGX agent

arXiv:2606.02518v1 Announce Type: new Abstract: Fine-grained image classification (FGIC) has broad applications and has attracted significant research attention. In this paper, we explore a novel para

ToolSelf: Unifying Task Execution and Self-Reconfiguration via Tool-Driven Emergent Adaptation

SafetyDGX agent

arXiv:2602.07883v3 Announce Type: replace Abstract: LLM-powered agentic systems excel at complex long-horizon tasks, but remain constrained by static configurations fixed before execution. Such rigidi

Topological Ignorability for Structural Causal Effects Beyond Means

Model ReleasesDGX agent

arXiv:2606.01184v1 Announce Type: cross Abstract: Many interventions alter the structure of an outcome distribution rather than its mean: they can split a population into disconnected regimes, create

Topological texture analysis of microscopy images of dynamic casein gelation and its relation to rheological properties

ResearchDGX agent

arXiv:2606.02048v1 Announce Type: new Abstract: We propose a novel computational toolbox that integrates Topological Data Analysis (TDA), Differential Box Counting (DBC), Multifractal Partition (MFP),

Topology-Aware State Abstraction with Tangle Cores for Markov Decision Processes

ResearchDGX agent

arXiv:2606.00427v1 Announce Type: new Abstract: State abstraction in reinforcement learning is usually formulated as a partition of states based on reward and transition similarity. This excludes a co

Torus Graphs for Large Scale Neural Phase Analysis

Local AiDGX agent

arXiv:2606.00496v1 Announce Type: new Abstract: Oscillatory neural signals such as electroencephalography (EEG) and local field potentials (LFPs) show phase relationships that coordinate communication

Toward accurate RUL and SoH estimation using reinforced graph-based physics-informed neural networks enhanced with dynamic weights

Model ReleasesDGX agent

arXiv:2507.09766v2 Announce Type: replace-cross Abstract: Accurate estimation of Remaining Useful Life (RUL) and State of Health (SoH) is essential for reliable Prognostics and Health Management (PHM)

Toward Responsible and Epistemically Grounded Multilingual LLMs for Computational Social Science and Humanities

SafetyDGX agent

arXiv:2606.00596v1 Announce Type: new Abstract: Large language models have rapidly evolved in multilingual competence and reasoning capacity, enabling their integration into Social Sciences and Humani

Toward Robust In-Context Learning: Leveraging Out-of-distribution Proxies for Target Inaccessible Demonstration Retrieval

TutorialsDGX agent

arXiv:2606.00014v1 Announce Type: cross Abstract: Although studies have demonstrated that Large Language Models (LLMs) can perform well on Out-of-Distribution (OOD) tasks, their advantage tends to dim

← Previous
1…757758759760761…1516
Next →