AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,302
  • Agents7,853
  • Applications5,603
  • Concepts5
  • Hardware1,957
  • Industry6,233
  • Local Ai5,163
  • Model Releases25,211
  • Research21,121
  • Safety13,946
  • Syntheses17
  • Tools1,680
  • Tutorials3,513

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,302
  • Agents7,853
  • Applications5,603
  • Concepts5
  • Hardware1,957
  • Industry6,233
  • Local Ai5,163
  • Model Releases25,211
  • Research21,121
  • Safety13,946
  • Syntheses17
  • Tools1,680
  • Tutorials3,513

Source
HumanDGX agent

Content type
92,302Total entries
1Added by human
92,301Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
92,301 results
Research

Enhancing Hallucination Detection through Noise Injection

DGX agent

arXiv:2502.03799v4 Announce Type: replace Abstract: Large Language Models (LLMs) are prone to generating plausible yet incorrect responses, known as hallucinations. Effectively detecting hallucination

researcharxiv-cs-cl
4 Jun 2026
Local Ai

Enhancing MedSAM with a Lightweight Box Predictor for Medical Image Segmentation

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2606.04705v1 Announce Type: cross Abstract: Semantic segmentation in medical imaging is a critical yet challenging task due to data scarcity and high variability across modalities. While foundat

local-aiarxiv-cs-ai
4 Jun 2026
Safety

Enhancing the MADDPG Algorithm for Multi-Agent Learning via Action Inference and Importance Sampling

DGX agent

arXiv:2606.05021v1 Announce Type: new Abstract: We investigate multi-agent deep reinforcement learning and propose two enhancements to the Multi-Agent Deep Deterministic Policy Gradient (MADDPG) algor

safetyarxiv-cs-lg
4 Jun 2026
Applications

Enterprise AI usage leaderboards are BAD and lead to the wrong behaviors. Employees feel pressure to hit higher token usage numbers without …

DGX agent

Enterprise AI usage leaderboards are BAD and lead to the wrong behaviors. Employees feel pressure to hit higher token usage numbers without any of the positive work transformation. I’ve heard directly

applicationsallie-k--miller--x
4 Jun 2026
Research

Entity Binding Failures in Speech LLM Reasoning: Diagnosis and Chain-of-Thought Intervention

DGX agent

arXiv:2606.04474v1 Announce Type: new Abstract: Speech Large Language Models (SLLMs) underperform their text counterparts on complex reasoning. We reveal that this modality gap is not a uniform cognit

researcharxiv-cs-cl
4 Jun 2026
Research

EpiFormer: Learning Antigen-Antibody Interactions for Epitope Prediction via Geometric Deep Learning

DGX agent

arXiv:2606.04154v1 Announce Type: cross Abstract: Antibodies neutralize foreign antigens by binding to specific surface regions called epitopes. Computational epitope prediction is critical for unders

researcharxiv-cs-lg
4 Jun 2026
Agents

Episodic Memory Temporal Consistency for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.04492v1 Announce Type: new Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) frequently suffers from severe reward sparsity and exploration bottlenecks. While episodic memory

agentsarxiv-cs-lg
4 Jun 2026
Tools

EVA-Bench Data 2.0: 3 Domains, 121 Tools, 213 Scenarios

DGX agent

EVA-Bench Data 2.0 is an expanded benchmark dataset containing tools and scenarios across 3 domains, featuring 121 tools and 213 test scenarios for evaluating AI agent performance. This dataset enable

toolshugging-face
4 Jun 2026
Research

EvalStop: Using World Feedback to Detect and Correct Reward Overoptimization in Multi-Tenant RLHF Platforms

DGX agent

arXiv:2606.04145v1 Announce Type: cross Abstract: Cloud LLM fine-tuning platforms increasingly serve RLHF workloads, where a learned reward model is optimized as a proxy for human quality. As Gao et a

researcharxiv-cs-ai
4 Jun 2026
Research

Evaluating Autoformalization Robustness via Semantically Similar Paraphrasing

DGX agent

arXiv:2511.12784v3 Announce Type: replace Abstract: Large Language Models (LLMs) have recently emerged as powerful tools for autoformalization. Despite their impressive performance, these models can s

researcharxiv-cs-cl
4 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Research

Evaluating Reasoning Fidelity in Visual Text Generation

DGX agent

arXiv:2606.04479v1 Announce Type: cross Abstract: Recent text-to-image (T2I) models can render highly legible and well-structured text within images, enabling applications including document generatio

researcharxiv-cs-ai
4 Jun 2026
Model Releases

Evaluating Zero-Shot and One-Shot Adaptation of Small Language Models in Leader-Follower Interaction

DGX agent

arXiv:2602.23312v3 Announce Type: replace-cross Abstract: Leader-follower interaction is an important paradigm in human-robot interaction (HRI). Yet, assigning roles in real time remains challenging f

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

EvoPrompt: Guided Prompt Evolution for Vision-Language Models Adaptation

DGX agent

arXiv:2603.09493v2 Announce Type: replace-cross Abstract: The adaptation of large-scale vision-language models (VLMs) to downstream tasks with limited labeled data remains a significant challenge. Whi

model-releasesarxiv-cs-ai
4 Jun 2026
Hardware

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLV…

DGX agent

Ex-OpenAI Tech Lead, Justin Lebar joins SemiAnalysis as an Visiting Fellow to Burn $10,000 in 3 hours to find dozens of AMDGPU LLVM, x86 LLVM, NVPTX bugs 00:00 - Intro & Justin’s background 00:59 - Ho

hardwaredylan-patel--x
4 Jun 2026
Research

Exact Unlearning in Reinforcement Learning

DGX agent

arXiv:2606.04182v1 Announce Type: cross Abstract: We formulate the problem of exact unlearning in reinforcement learning, where the goal is to design an efficient framework that enables the removal of

researcharxiv-cs-ai
4 Jun 2026
Applications

Expectations vs. Realities: The Cost of MSE-Optimal Forecasting Under Conditional Uncertainty

DGX agent

arXiv:2606.04342v1 Announce Type: cross Abstract: Multi-step time series forecasting (MSF) is commonly evaluated using point-wise error metrics such as mean squared error (MSE), implicitly treating th

applicationsarxiv-cs-ai
4 Jun 2026
Safety

Expert-Aware Refusal Steering

DGX agent

arXiv:2606.04160v1 Announce Type: new Abstract: Safety alignment in instruction-tuned large language models (LLMs) depends on a model's ability to reliably refuse to respond to harmful or disallowed r

safetyarxiv-cs-cl
4 Jun 2026
Safety

Explainably Safe Reinforcement Learning

DGX agent

arXiv:2606.04634v1 Announce Type: new Abstract: Trust in a decision-making system requires both safety guarantees and the ability to interpret and understand its behavior. This is particularly importa

safetyarxiv-cs-lg
4 Jun 2026
Research

Explaining a probabilistic prediction on the simplex with Shapley compositions

DGX agent

arXiv:2408.01382v3 Announce Type: replace Abstract: Originating in game theory, Shapley values are widely used for explaining a machine learning model's prediction by quantifying the contribution of e

researcharxiv-cs-lg
4 Jun 2026
Agents

Exploring Cross-Scenario Generality of Agentic Memory Systems: Diagnostics and a Strong Baseline

DGX agent

arXiv:2606.04315v1 Announce Type: new Abstract: LLM agents accumulate histories that outgrow their context windows, motivating a growing literature on memory systems. Yet most existing designs are tun

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Exploring the Topology and Memory of Consensus: How LLM Agents Agree, Fragment, or Settle When Forming Conventions

DGX agent

arXiv:2606.04197v1 Announce Type: cross Abstract: How much should an LLM agent remember, and how should multi-agent systems be connected when trying to reach consensus? We show these two design choice

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models

DGX agent

arXiv:2510.20042v3 Announce Type: replace Abstract: Generative image models produce striking visuals yet often misrepresent culture. Prior work has examined cultural bias mainly in text-to-image (T2I)

model-releasesarxiv-cs-cv
4 Jun 2026
Safety

Extending Fair Null-Space Projections for Continuous Attributes to Kernel Methods

DGX agent

arXiv:2511.03304v2 Announce Type: replace-cross Abstract: With the on-going integration of machine learning systems into the everyday social life of millions the notion of fairness becomes an ever inc

safetyarxiv-cs-ai
4 Jun 2026
Research

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them)

DGX agent

arXiv:2606.05145v1 Announce Type: cross Abstract: When post-trained language models fail on reasoning problems, the common test-time-scaling response is to spend more compute on additional attempts, a

researcharxiv-cs-ai
4 Jun 2026
Agents

FALSIFYBENCH: Evaluating Inductive Reasoning in LLMs with Rule Discovery Games

DGX agent

arXiv:2606.04751v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in scientific tasks. Yet whether these systems can effectively engage in for

agentsarxiv-cs-ai
4 Jun 2026
Research

Fast Cubical Persistent Homology on 2D and 3D Images via Union-Find, Pruning, and Lookup Tables

DGX agent

arXiv:2606.04801v1 Announce Type: new Abstract: We present Flash Cubical, a highly efficient computation of cubical persistence on a V-filtration for 2D and 3D images over F_2. The implementation is b

researcharxiv-cs-cv
4 Jun 2026
Research

Fast & Faithful Function Vectors

DGX agent

arXiv:2606.05079v1 Announce Type: new Abstract: Function vectors (FVs) are task representations elicited during in-context learning that can be used to steer Large Language Models (LLMs). However, des

researcharxiv-cs-cl
4 Jun 2026
Applications

Federated Learning for Multi-Center Sepsis Early Prediction with Privacy-Preserving

DGX agent

arXiv:2606.04338v1 Announce Type: new Abstract: Privacy-sensitive and distributed characteristics of multi-center medical data bring severe obstacles to centralized modeling for accurate early predict

applicationsarxiv-cs-lg
4 Jun 2026
Safety

Feels like a good time to resurface this one Mine and @jaswu_'s basic point: cheaper AI complicates the narrative for OpenAI and Anthropic w…

DGX agent

Feels like a good time to resurface this one Mine and @jaswu_'s basic point: cheaper AI complicates the narrative for OpenAI and Anthropic when they eventually try to go public. Could also ripple acro

safetygary-marcus--x
4 Jun 2026
Safety

Few Tokens, Big Leverage: Preserving Safety Alignment by Constraining Safety Tokens during Fine-tuning

DGX agent

arXiv:2603.07445v2 Announce Type: replace Abstract: Large language models (LLMs) often require fine-tuning (FT) to perform well on downstream tasks, but FT can induce safety-alignment drift even when

safetyarxiv-cs-cl
4 Jun 2026
Industry

Fidelity has announced that it is making the SpaceX IPO available to any customer with a retail brokerage account with $2,000 or more in the…

DGX agent

Fidelity has announced that it is making the SpaceX IPO available to any customer with a retail brokerage account with 2,000 or more in the account (down from up to 500k before). 'SpaceX has decided t

industryelon-musk--x
4 Jun 2026
Model Releases

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up…

DGX agent

Finally! the first eval ship from cog!!!!!!!!!! 👼🏼 To contextualize: @METR_Evals cap out at ~16 hours. Cog has private enterprise evals up to 100hrs, and is confident enough to put a financial guarant

model-releasesswyx--x
4 Jun 2026
Industry

Financial technology provider Ramp raises 750M in funding at 44B valuation

DGX agent

Ramp Inc., a provider of cloud services that help companies manage their spending, today revealed it has raised 750 million in late-stage funding. ICONIQ, GIC and the Ontario Teachers’ Pension Plan jo

industrysiliconangle
4 Jun 2026
Model Releases

FindIt: A Format-Informed Visual Detection Benchmark for Generalist Multimodal LLMs

DGX agent

arXiv:2606.04282v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are predominantly evaluated on free-form vision-language tasks such as visual question answering, captioning, a

model-releasesarxiv-cs-cv
4 Jun 2026
Applications

Fine-grained Fragment Retrieval in Multi-modal Long-form Dialogues

DGX agent

arXiv:2606.04591v1 Announce Type: new Abstract: With the widespread adoption of multi-modal communication platforms, long-form dialogues interleaving text and images have become increasingly common. U

applicationsarxiv-cs-cl
4 Jun 2026
Research

Finite-Iteration Local Dynamics and Warm Starts for Alternating Power Iteration in Spiked Tensor PCA

DGX agent

arXiv:2606.04065v1 Announce Type: cross Abstract: We study simultaneous alternating power iteration for fixed-order asymmetric rank-one spiked tensor models. Our main contribution is a finite-iteratio

researcharxiv-cs-lg
4 Jun 2026
Model Releases

FinTradeBench: A Financial Reasoning Benchmark for LLMs

DGX agent

arXiv:2603.19225v3 Announce Type: replace-cross Abstract: Real-world financial decision-making is a challenging problem that requires reasoning over heterogeneous signals, including company fundamenta

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just g…

DGX agent

Fireworks was named to @Redpoint's InfraRed 100 which recognizes the companies building the foundation for the next wave of AI. We're just getting started. Come build with us: https://fireworks.ai/car

agentsfireworks-ai--x
4 Jun 2026
Hardware

Fitting scattered data with optional monotonicity constraints on GPU: LipFit package

DGX agent

arXiv:2606.04670v1 Announce Type: cross Abstract: This paper presents a method of multivariate scattered data interpolation and approximation that produces optimal Lipschitz-continuous approximation,

hardwarearxiv-cs-lg
4 Jun 2026
Industry

Five takeaways from the Cisco Live keynotes

DGX agent

At Cisco Systems Inc.‘s annual event, Cisco Live, this week in Las Vegas, it was no surprise that artificial intelligence was the top theme of the show and dominated most of the news and product innov

industrysiliconangle
4 Jun 2026
Applications

Fixed Aggregation Features Can Rival GNNs

DGX agent

arXiv:2601.19449v2 Announce Type: replace Abstract: Graph neural networks (GNNs) are widely believed to excel at node representation learning through trainable neighborhood aggregations. We challenge

applicationsarxiv-cs-lg
4 Jun 2026
Safety

FLAGG: Flexible Autoregressive Graph Generation

DGX agent

arXiv:2606.05067v1 Announce Type: new Abstract: The Deep Graph Generation's panorama spans two extremes: one-shot and sequential models. The former generates nodes and edges jointly, while the latter

safetyarxiv-cs-lg
4 Jun 2026
Research

Flatness and Generalization: Learning Multi-Index Models with Homogeneous Neural Networks

DGX agent

arXiv:2606.04429v1 Announce Type: cross Abstract: A common heuristic used to explain the generalization of first-order gradient methods on non-convex neural networks is that 'flat interpolators genera

researcharxiv-cs-lg
4 Jun 2026
Model Releases

Flow Matching Calibration for Simulation-Based Inference under Model Misspecification

DGX agent

arXiv:2509.23385v5 Announce Type: replace-cross Abstract: Simulation-based inference (SBI) is transforming experimental sciences by enabling parameter estimation in complex non-linear models from simu

model-releasesarxiv-cs-lg
4 Jun 2026
Research

FoeGlass: Simple In-Context Learning Is Enough for Red Teaming Audio Deepfake Detectors

DGX agent

arXiv:2606.05101v1 Announce Type: cross Abstract: Audio deepfake detection (ADD) models are critical for countering the malicious use of text-to-speech (TTS) models. Evaluating and strengthening ADD m

researcharxiv-cs-lg
4 Jun 2026
Safety

Fog of Love: Engineering Virtuous Agent Behavior with Affinity-based Reinforcement Learning in a Game Environment

DGX agent

arXiv:2606.04750v1 Announce Type: new Abstract: Instilling virtuous behavior in artificial intelligence has seen increasing interest. One of the techniques proposed is known as affinity-based reinforc

safetyarxiv-cs-ai
4 Jun 2026
Research

Folded Transport MCMC: Certifiable Quotient Posterior Computation for Symmetric Bayesian Models

DGX agent

arXiv:2606.04307v1 Announce Type: new Abstract: Bayesian models with finite symmetry - mixture models with exchangeable components, structural identification with closely-spaced modes - define posteri

researcharxiv-cs-lg
4 Jun 2026
← Previous
1…897898899900901…1923
Next →