AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
Human
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
91,059 results
9 Jun 2026

Symbolic Reasoning Frameworks Modulate LLM Risk Aversion in Multi-Agent Strategic Settings

SafetyDGX agent

arXiv:2606.07552v1 Announce Type: cross Abstract: Large language models exhibit innate behavioral tendencies when deployed as strategic agents -- notably a risk-averse 'turtle' bias toward defensive p

Symskill: Symbol and Skill Co-Invention for Data-Efficient and Reactive Long-Horizon Manipulation

ResearchDGX agent

arXiv:2510.01661v3 Announce Type: replace Abstract: Multi-step manipulation in dynamic environments remains challenging. Imitation learning (IL) is reactive but lacks compositional generalization, sin

SynManDex: Synthesizing Human-like Dexterous Grasps from Synthetic Human Pre-Grasps

ResearchDGX agent

arXiv:2606.09798v1 Announce Type: new Abstract: Human hand-object interactions encode functional intent, but direct transfer to robotic hands often fails under morphology, contact, and reachability co

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Synthetic but Not Realistic: The Evaluation Challenge in Generative Modelling for Structured Electronic Medical Records

ApplicationsDGX agent

arXiv:2606.08903v1 Announce Type: new Abstract: Synthetic healthcare data are widely proposed as privacy-preserving substitutes for real patient data, yet their evaluation remains dominated by statist

SynthICL: Scalable In-context Imitation Learning with Synthetic Data

SafetyDGX agent

arXiv:2606.08154v1 Announce Type: new Abstract: In-context imitation learning (ICIL) enables robots to learn new tasks from a small number of demonstrations by conditioning a pre-trained policy on tas

Systematic LLM Translation of Legacy Scientific Code to Differentiable Frameworks: Application to a Land Surface Model

Model ReleasesDGX agent

arXiv:2606.07681v1 Announce Type: cross Abstract: Differentiable programming offers transformative capabilities for scientific modeling, enabling gradient-based parameter estimation, sensitivity analy

Systems-Level Planning and Coordination of Truck-Drone Collaborative Delivery Networks

SafetyDGX agent

arXiv:2606.08738v1 Announce Type: cross Abstract: Urban last-mile parcel delivery increasingly relies on heterogeneous fleets whose performance depends on timely coordination, reliable communication,

TABVERSE: Benchmarking Cross-Format Table Understanding in LLMs and VLMs

Model ReleasesDGX agent

arXiv:2606.09578v1 Announce Type: new Abstract: Large Language Models (LLMs) and Vision-Language Models (VLMs) are increasingly evaluated on table reasoning tasks, but the role of table representation

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

Model ReleasesDGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

Taming Perception Jitter: Uncertainty-Aware LiDAR Object Detection for Reliable Motion Classification

AgentsDGX agent

arXiv:2606.09350v1 Announce Type: cross Abstract: Reliable motion classification is critical for autonomous driving, as false dynamic predictions of static objects can cascade into unnecessary planner

TAMUNA: Doubly Accelerated Distributed Optimization under Partial Participation

Local AiDGX agent

arXiv:2302.09832v4 Announce Type: replace Abstract: In distributed optimization and federated learning, slow and costly communication between parallel devices and the central server constitutes the pr

TAO: Tolerance-Aware Optimistic Verification for Floating-Point Neural Networks

HardwareDGX agent

arXiv:2510.16028v4 Announce Type: replace-cross Abstract: Neural networks increasingly run on hardware outside the user's control (cloud GPUs, inference marketplaces). Yet ML-as-a-Service reveals litt

Targeting World Models to Compromise Robot Learning Pipelines

SafetyDGX agent

arXiv:2606.09499v1 Announce Type: cross Abstract: World models have recently seen a rapid growth in both their popularity and capability as more data efficient tools for generating robot training data

TBD-VLA: Temporal Block Diffusion Vision Language Action Model

ApplicationsDGX agent

arXiv:2606.07895v1 Announce Type: new Abstract: Discrete Vision-Language-Action (VLA) models typically formulate action generation as next-token prediction over discretized action spaces, conditioning

Teacher-Free Self-Training Amplifies but Does Not Compound: A Pass@K Crossover on a Free-Verifier Domain

Local AiDGX agent

arXiv:2606.07856v1 Announce Type: new Abstract: When a language model trains on its own verified outputs, does it acquire capability beyond its base, or merely get better at expressing capability the

TeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning

ResearchDGX agent

arXiv:2606.08770v1 Announce Type: cross Abstract: The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes

@tejalpatwardhan this directly inspired my work at @cognition frontiercode https://x.com/jai_torregrosa/status/2064100437549072593?s=20

ToolsDGX agent

@tejalpatwardhan this directly inspired my work at @cognition frontiercode https://x.com/jai_torregrosa/status/2064100437549072593?s=20 props for including your own model on the chart. it's making a t

TempoBench: Evaluating Temporal Causal Reasoning in Large Language Models

Model ReleasesDGX agent

arXiv:2510.27544v2 Announce Type: replace Abstract: Temporal reasoning involves understanding how systems evolve over time through input-driven state transitions. A key aspect is temporal causal reaso

Temporal-Aware Reasoning Optimization for Video Temporal Grounding

Local AiDGX agent

arXiv:2606.09248v1 Announce Type: new Abstract: Multi-modal Large Language Models (MLLMs) have achieved remarkable progress in video temporal grounding with reinforcement learning for generating reaso

Temporal Coverage over Density: Parsimonious Training-Set Design for ML Climate Downscaling

ResearchDGX agent

arXiv:2606.07898v1 Announce Type: new Abstract: High-resolution regional climate simulations provide critical information for climate impacts assessments but remain computationally expensive, motivati

Tensorizing Engram: Sharing Latents Across N-Gram Embeddings is Beneficial in LLMs

ResearchDGX agent

arXiv:2606.08347v1 Announce Type: cross Abstract: Modern language models represent text using discrete token-level embeddings, which forces recurring multi-token patterns to be learned implicitly acro

Tesla AI chip design engineering reviews are so great! Team is awesome. Our AI6 chip might set a record for most amount of usable intelligen…

IndustryDGX agent

Elon Musk posted on X about Tesla's AI chip design engineering team and their work on the AI6 chip, expressing confidence in the team's performance and suggesting the chip could achieve notable capabi

Test-Time Adaptive Composition for Machine Learning as a Service (MLaaS) in IoT Environments

ResearchDGX agent

arXiv:2606.07685v1 Announce Type: cross Abstract: The dynamic nature of Internet of Things (IoT) environments affects the long-term effectiveness of Machine Learning as a Service (MLaaS) compositions.

Test-Time Scaling in Multimodal Foundation Models: A Comprehensive Survey of Generation and Reasoning

ResearchDGX agent

arXiv:2606.08231v1 Announce Type: new Abstract: Test-time Scaling (TTS) has emerged as a pivotal research direction for enhancing model performance by dynamically allocating computational resources du

Testing the Black Box: Structural Barriers to Independent Evaluation of Consumer-Facing Health LLMs

SafetyDGX agent

arXiv:2606.08483v1 Announce Type: new Abstract: Background: Consumer-facing large language models are now a common source of health information, and they interpret and personalize responses rather tha

The ACUTE Protocol: Operationalizing Language Model Activations for Better Calibration, Utility, and Trust

SafetyDGX agent

arXiv:2606.07822v1 Announce Type: cross Abstract: As language models improve and become increasingly deployed to solve a variety of tasks, trustworthiness becomes essential. Calibration is a good prox

The AI Epistemic Deference Index: A Continuous Measure of Sycophancy

Model ReleasesDGX agent

arXiv:2606.07897v1 Announce Type: new Abstract: Current AI models frequently exhibit epistemic sycophancy, endorsing claims to agree with a user. Existing evaluations typically measure this either by

The AI supersystem shift: Why Arista’s 1.6T announcement is an Ethernet inflection point

IndustryDGX agent

The networking industry loves inflection points. Over the years, we have had many new compute models that require the network to evolve. For as long as I can remember, the holy war between InfiniBand

The best AI infrastructure shouldn't be reserved for the biggest companies. Together AI is partnering with @pax8 to bring powerful, cost-eff…

ApplicationsDGX agent

The best AI infrastructure shouldn't be reserved for the biggest companies. Together AI is partnering with @pax8 to bring powerful, cost-efficient AI and leading open-source models to small and mid-si

The CIFAR Synthetic Evidence Corpus for Detecting AI-Generated Evidence

TutorialsDGX agent

arXiv:2606.07916v1 Announce Type: new Abstract: The growing ability of generative models to produce realistic documents poses a direct challenge to evidentiary workflows in the justice system and the

The Confidence Trap: Calibration Attacks for Graph Neural Networks

SafetyDGX agent

arXiv:2606.08467v1 Announce Type: cross Abstract: While confidence calibration is essential for trustworthy decision-making in safety-critical applications, the robustness of calibrated GNNs to advers

// The Consistency Illusion // Multi-agent debate can make agents agree on the final answer while their underlying reasoning stays misaligne…

AgentsDGX agent

// The Consistency Illusion // Multi-agent debate can make agents agree on the final answer while their underlying reasoning stays misaligned. This work finds that consensus on the output hides disagr

The Cross-Architecture Substrate: A Domain-Transcendent, Calibration-Surviving Geometric Invariant of Modern Vision Encoders

SafetyDGX agent

arXiv:2606.07882v1 Announce Type: cross Abstract: Different vision neural networks -- trained to classify, contrast, reconstruct, or match images to text -- should have correspondingly different inter

The Download: whole-body rejuvenation drugs and five things to know about AI

ResearchDGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. David Sinclair plans to test whole-body rejuvenation drugs in

The Easy, the Hard, and the Learnable: Confidence and Difficulty-Adaptive Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.07950v1 Announce Type: new Abstract: RL with verifiable rewards can substantially improve LLM reasoning, yet standard GRPO-style training often treats easy, hard, and learnable questions al

The EU warns that AI-boosted chemical synthesis is helping European drug gangs develop new 'designer' drug precursors that evade existing product blacklists (Michael Peel/Financial Times)

IndustryDGX agent

Michael Peel / Financial Times: The EU warns that AI-boosted chemical synthesis is helping European drug gangs develop new “designer” drug precursors that evade existing product blacklists — EU agency

The evidence of two tier policing and judiciary is clear for all to see.

IndustryDGX agent

This post claims that there is evident disparity in how the police and judicial system treat different groups or individuals. The assertion suggests systemic inconsistency in law enforcement and court

The fact that Anthropic may take away subscription access to Fable in two weeks is weird & discourages investing in learning about the model…

ApplicationsDGX agent

The fact that Anthropic may take away subscription access to Fable in two weeks is weird & discourages investing in learning about the model. Subscription use is how you figure out what the model is g

The Flexibility Trap: Rethinking the Value of Arbitrary Order in Diffusion Language Models

SafetyDGX agent

arXiv:2601.15165v4 Announce Type: replace-cross Abstract: Diffusion Large Language Models (dLLMs) break the rigid left-to-right constraint of traditional LLMs, enabling token generation in arbitrary o

The Governance of Human-LLM Interaction: Safety Gating, Civility Steering, and Affective Default Lock-In

SafetyDGX agent

arXiv:2606.08172v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate high-stakes interactions in finance, medicine, and mental-health support, yet users have limited con

The Hidden Bias of Process Reward Models:PRISM for Rewarding the Right Reasoning

SafetyDGX agent

arXiv:2606.09078v1 Announce Type: new Abstract: Process Reward Models (PRMs) improve credit assignment for reasoning by providing step-level feedback. However, we identify a hidden bias in PRMs caused

The Injection Paradox: Brand-Level Suppression in Safety-Trained LLM Recommendations via RAG Context Injection

Model ReleasesDGX agent

arXiv:2606.09204v1 Announce Type: new Abstract: We present a reproducible failure mode of safety training in RAG-based LLM recommendation -- the Injection Paradox -- in which prompt injections embedde

The Label Horizon Paradox: Rethinking Supervision Targets in Financial Forecasting

ResearchDGX agent

arXiv:2602.03395v4 Announce Type: replace Abstract: While deep learning has revolutionized financial forecasting through sophisticated architectures, the design of the supervision signal itself is rar

The Last Visible Pixel: Probing Fine-Scale Perception in Vision-Language Models

Model ReleasesDGX agent

arXiv:2606.07861v1 Announce Type: cross Abstract: Recent vision-language models (VLMs) excel at multimodal understanding and reasoning, yet their fine-grained visual perception remains underexplored.

The Mirrored Influence Hypothesis: Efficient Data Influence Estimation by Harnessing Forward Passes

ResearchDGX agent

arXiv:2402.08922v3 Announce Type: replace Abstract: Large-scale black-box models have become ubiquitous across numerous applications. Understanding the influence of individual training data sources on

The Montparnasse Algorithm for RNA Design

Model ReleasesDGX agent

arXiv:2606.07562v1 Announce Type: cross Abstract: RNA design consists of discovering a nucleotide sequence that optimizes predefined criteria, such as secondary structure. It is useful for synthetic b

The Need for Neural ISP in the Small-Pixel Era: How Shrinking Pixels Push Optics to the Limit and Neural Restoration Pushes Back

Local AiDGX agent

arXiv:2606.07675v1 Announce Type: cross Abstract: Smartphone telephoto cameras are approaching a 'telephoto physics wall': as pixel pitches shrink toward sub-0.5 micron, the optics remain limited by g

The New York Times published a roundtable discussion between @DAcemogluMIT, @deanwball, @clarashih & myself about the future of AI & who win…

ApplicationsDGX agent

The New York Times published a roundtable discussion between @DAcemogluMIT, @deanwball, @clarashih & myself about the future of AI & who wins at work. I think it is a really nice overview of the core

The Routing Plateau: Understanding and Breaking the Accuracy Limits of LLM Routers

TutorialsDGX agent

arXiv:2606.07587v1 Announce Type: new Abstract: LLM routing has become a popular approach to improve the cost-quality trade-off of LLM services by dynamically selecting a model for each query. Recent

The Sample Complexity of Parameter-Free Stochastic Convex Optimization

Model ReleasesDGX agent

arXiv:2506.11336v2 Announce Type: replace Abstract: We study the sample complexity of stochastic convex optimization when problem parameters such as the distance to optimality and the Lipschitz consta

The Spectral Dynamics and Noise Geometry of Muon

SafetyDGX agent

arXiv:2606.08388v1 Announce Type: new Abstract: Muon replaces a matrix gradient G=USigma V^op by its polar factor UV^op. This keeps the singular directions selected by the gradient, but makes the upda

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

AgentsDGX agent

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

The Topological Dual of a Dataset: A Logic-to-Topology Encoding for AlphaGeometry-Style Data

ResearchDGX agent

arXiv:2604.18050v2 Announce Type: replace Abstract: AlphaGeometry represents a milestone in neuro-symbolic reasoning, yet its architecture faces a log-linear scaling bottleneck within its symbolic ded

The truth is that there are VASTLY more hate crimes, especially aggravated rape and murder, per person by Blacks against Whites than the oth…

IndustryDGX agent

The truth is that there are VASTLY more hate crimes, especially aggravated rape and murder, per person by Blacks against Whites than the other way around. The is not remotely debatable, as the numbers

The US FCC waives its deadline for Amazon to deploy half of its Leo satellites by July; Amazon is still required to launch all 3,232 satellites by July 30, 2029 (Michael Kan/PCMag)

ApplicationsDGX agent

Michael Kan / PCMag: The US FCC waives its deadline for Amazon to deploy half of its Leo satellites by July; Amazon is still required to launch all 3,232 satellites by July 30, 2029 — The Federal Comm

The Value of Personalized Recommendations: Evidence from Netflix

ResearchDGX agent

arXiv:2511.07280v5 Announce Type: replace-cross Abstract: Personalized recommendation systems shape much of user choice online, yet their targeted nature makes separating out the value of recommendati

TheoremBench: Evaluating LLMs on Theorem Proving in Formal Mathematics

Model ReleasesDGX agent

arXiv:2606.09450v1 Announce Type: new Abstract: LLMs have recently achieved strong results on formal proving benchmarks. However, existing evaluations remain heavily concentrated on competition-style

Theoretical Foundations of Continual Learning via Drift-Plus-Penalty

Model ReleasesDGX agent

arXiv:2606.08452v1 Announce Type: new Abstract: In many real-world settings, data streams are nonstationary and arrive sequentially, requiring learning systems to adapt continuously without retraining

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at h…

Model ReleasesDGX agent

There has been a lot of hand wringing on the appropriate valuation of SpaceX. Some large institutions believe SpaceX can only be valued at half what the market seems to be willing to pay for it. Other

They call for calm in the face of horror. This is inhuman and manipulative. The proportionate and natural response to an horrific incident i…

IndustryDGX agent

They call for calm in the face of horror. This is inhuman and manipulative. The proportionate and natural response to an horrific incident is disgust, fear and anger. They want to downplay the horror

← Previous
1…655656657658659…1518
Next →