AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
Human
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
90,315 results
29 May 2026

Improved Guarantees for Heterogeneous Treatment-Effect Estimation via Matrix Completion

ResearchDGX agent

arXiv:2605.30319v1 Announce Type: cross Abstract: A central goal of modern causal inference is estimating heterogeneous treatment effects to answer questions like 'how does an intervention affect each

Improving Adversarial Robustness of Attribution via Implicit Regularization

Model ReleasesDGX agent

arXiv:2605.29983v1 Announce Type: cross Abstract: The adversarial robustness of attributions is a fundamental requirement for reliable explainability in deep learning, yet existing approaches typicall

Improving agents The old way: Manually reading traces, looking for patterns, writing evals, and creating fixes. The better way: Letting Lang…

AgentsDGX agent
DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

This post from LangChain's Harrison Chase contrasts traditional manual methods of improving AI agents (tracing execution, identifying patterns, writing evaluations, and implementing fixes) with a more

Improving CLIP Adaptation by Breaking Tail Alignment for Source-Free Cross-Domain Few-Shot Learning

SafetyDGX agent

arXiv:2605.29776v1 Announce Type: new Abstract: Vision-Language Models (VLMs) such as CLIP demonstrate strong zero-shot generalization, but their performance significantly degrades in cross-domain sce

Improving Collaborative Storytelling with a Multi-Agent Framework Based on Large Language Models

AgentsDGX agent

arXiv:2605.29625v1 Announce Type: new Abstract: The topic of Co-creation, i.e., AI agents interacting with humans to generate outputs (e.g., art), has gained significant attention recently. However, m

Improving Full Waveform Inversion in Large Model Era

Model ReleasesDGX agent

arXiv:2603.00377v2 Announce Type: replace Abstract: Full Waveform Inversion (FWI) is a highly nonlinear and ill-posed problem that aims to recover subsurface velocity maps from surface-recorded seismi

In case anyone wants to improve/change/use it: https://github.com/emollick/veil-of-history

ApplicationsDGX agent

The Veil of History is an open-source project shared by Ethan Mollick on GitHub that appears to be a tool or resource related to historical analysis or visualization, made available for public improve

In-Context Reward Adaptation for Robust Preference Modeling

SafetyDGX agent

arXiv:2605.30323v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) typically relies on static reward models to align Large Language Models with human preferences. Howe

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve …

Model ReleasesDGX agent

In conversation with OpenAI’s @markchen90, Terence reflects on a future where AI reduces the cognitive friction of research, helps preserve the paths behind discovery, and expands what mathematicians

In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration

ResearchDGX agent

arXiv:2510.00777v2 Announce Type: replace Abstract: LLM-generated drafts often contain subtle factual or logical errors, yet prior work shows that models struggle to reliably integrate multi-turn feed

In the UK, you can stab a white person to death and get them arrested while they bleed to death, because racism. In the UK, you can brutally…

IndustryDGX agent

In the UK, you can stab a white person to death and get them arrested while they bleed to death, because racism. In the UK, you can brutally assault police officers on video and go free, because racis

Indexing the Unreadable: LLM-Native Recursive Construction and Search of Service Taxonomies

Model ReleasesDGX agent

arXiv:2605.29270v1 Announce Type: new Abstract: The era of the Internet of Agents (IoA) is taking shape: LLM agents are expected to fulfill user goals by orchestrating fast-growing populations of Mode

Inferring Code Correctness from Specification

SafetyDGX agent

arXiv:2605.29822v1 Announce Type: cross Abstract: Large language models (LLMs) have become integral to modern software development, enabling automated code generation at scale. However, validating the

Inferring the Size of Large Language Models From Popular Text Memorization

Model ReleasesDGX agent

arXiv:2605.29223v1 Announce Type: new Abstract: The parameter counts of the most widely used large language models (LLMs) are often withheld by their developers, leaving model size -- a primary refere

Influence-Guided Symbolic Regression: Scientific Discovery via LLM-Driven Equation Search with Granular Feedback

TutorialsDGX agent

arXiv:2605.29184v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a promising avenue for scientific discovery, yet their application to symbolic regression is often constrained by i

Inform, Coach, Relate, Listen: Auditing LLM Caregiving Support Roles

Model ReleasesDGX agent

arXiv:2605.29473v1 Announce Type: cross Abstract: Language models are increasingly being deployed for conversational support in informal caregiving contexts, where interactions often extend beyond inf

Information-Directed Offline-to-Online Reinforcement Learning

SafetyDGX agent

arXiv:2605.29405v1 Announce Type: new Abstract: Decision-making from offline datasets typically warm-starts a policy or score model from fixed offline data and then refines it with limited online inte

InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents

Model ReleasesDGX agent

arXiv:2511.22884v2 Announce Type: replace Abstract: Data analysis has become an indispensable part of scientific research. To discover the latent knowledge and insights hidden within massive datasets,

Inspectorch: Efficient rare event exploration in solar observations

ResearchDGX agent

arXiv:2602.20316v2 Announce Type: replace-cross Abstract: The Sun is observed in unprecedented detail, enabling studies of its activity on very small spatiotemporal scales. However, the large volume o

Instance-dependent Stochastic Lipschitz bandit

ResearchDGX agent

arXiv:2605.29748v1 Announce Type: cross Abstract: We study the Lipschitz bandit problem, where a learner sequentially maximizes an unknown Lipschitz function f over a domain X subset [0,1]^d using noi

Intent-aligned Autonomous Spacecraft Guidance via Reasoning Models

SafetyDGX agent

arXiv:2604.17176v2 Announce Type: replace-cross Abstract: Future spacecraft operations require autonomy that can interpret high-level mission intent while preserving safety. However, existing trajecto

Interactive In-Meeting Speaker Correction with Human Feedback

ResearchDGX agent

arXiv:2509.18377v2 Announce Type: replace Abstract: Most automatic speech processing systems operate in ``open loop'' mode without user feedback about who said what, yet human-in-the-loop workflows ca

Interesting that the GPT-5 Pro series models have consistently been the best models for single-shot attempts at the hardest problems since l…

Model ReleasesDGX agent

Interesting that the GPT-5 Pro series models have consistently been the best models for single-shot attempts at the hardest problems since last summer. There has been no real competition in all that t

Internal memo: Meta plans to start testing an AI pendant in 2027, release new AI glasses next month, start a 'Wearables for Work' unit for enterprises, and more (Jyoti Mann/The Information)

IndustryDGX agent

Jyoti Mann / The Information: Internal memo: Meta plans to start testing an AI pendant in 2027, release new AI glasses next month, start a “Wearables for Work” unit for enterprises, and more — Meta Pl

Internal Representation, Not Clinical Knowledge: Where Apparent LLM Triage Failures Originate

Model ReleasesDGX agent

arXiv:2605.29889v1 Announce Type: cross Abstract: Patient-voiced clinical-triage benchmarks report high under-triage rates for consumer LLMs for constrained multiple-choice output, yet the same cases

IP-Adapter Is All You Need: Towards Fine-Tuning-Free Diffusion-Based Talking Face Generation

ResearchDGX agent

arXiv:2605.30230v1 Announce Type: new Abstract: With the rapid advancement of diffusion models, talking face generation has made remarkable progress. However, existing diffusion-based methods still re

Is Your Diffusion Sampler Actually Correct? A Sampler-Centric Evaluation of Discrete Diffusion Language Models

ResearchDGX agent

arXiv:2602.19619v2 Announce Type: replace Abstract: Discrete diffusion language models (dLLMs) provide a fast and flexible alternative to autoregressive models (ARMs) via iterative denoising with para

Is Your LLM Overcharging You? Tokenization, Transparency, and Incentives

Model ReleasesDGX agent

arXiv:2505.21627v4 Announce Type: replace-cross Abstract: State-of-the-art large language models require specialized hardware and substantial energy to operate. As a consequence, cloud-based services

It might be time to fill one of these out again. Please remit to myself or @GaryMarcus Thank you.

SafetyDGX agent

Gary Marcus is requesting that someone complete a form or document and submit it to him or another person (possibly Gary Marcus himself based on the mention of @GaryMarcus). The post appears to be a r

It’s a good day when the Pope vouches for your recent comment in Nature.

SafetyDGX agent

It’s a good day when the Pope vouches for your recent comment in Nature. The Pope is making exactly our point. LLMs “may imitate or even simulate, but they do not understand.” This is the core epistem

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too m…

SafetyDGX agent

It's a shame what happened to @kevinroose. He used to be a pretty damn good tech reporter. But recently he got one-shotted by spending too much time in/around the big AI Labs (research for the book he

It`s All About Speed: AI`s Impact on Workflow in Music Production

ApplicationsDGX agent

arXiv:2605.29931v1 Announce Type: new Abstract: In this paper, we present the results of an ethnographic study into the impact of AI and automated tools on music production workflow. Focusing specific

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my …

AgentsDGX agent

I've been using state-of-the-art models to teach small models running on my computer how I work. The result : a personal agent that runs my inbox, my deal pipeline, my blog, my calendar, & my research

I've largely switched over to using GPT-5.5 in recent weeks, which I like nearly as much as Opus 4.6 and 4.7, and is *very* reasonably price…

Model ReleasesDGX agent

Jeremy Howard expresses positive views on GPT-5.5, stating he has recently switched to using it as his primary model and finds it nearly comparable to Anthropic's Opus 4.6 and 4.7 while offering signi

Jailbreaking and Mitigation of Vulnerabilities in Large Language Models

SafetyDGX agent

arXiv:2410.15236v4 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have transformed artificial intelligence by advancing natural language understanding and generation, enabling app

Jensen Huang called Fireworks 'the TSMC of AI factories' at GTC 2026. Here's the @nvidia CEO's full conversation with our own, @lqiao:

HardwareDGX agent

Jensen Huang, NVIDIA's CEO, compared Fireworks to 'the TSMC of AI factories' during a conversation at GTC 2026, highlighting Fireworks' role as a foundational infrastructure provider in the AI industr

@jerryjliu0 We also automatically update the ParseBench leaderboard on Kaggle :) https://www.kaggle.com/benchmarks/llamaindex-org/parsebench

AgentsDGX agent

Jerry Liu announces that the ParseBench leaderboard is automatically updated on Kaggle, providing a continuously maintained benchmark for parsing performance metrics. The leaderboard is hosted under t

Joint Angle Estimation with Customized Wristband Based on Online Incremental Learning

ResearchDGX agent

arXiv:2605.29771v1 Announce Type: new Abstract: Intelligent wearable technology plays an increasingly important role in human-computer interaction, motion, and health monitoring. To ensure comfort and

Joint Model and Data Sparsification via the Marginal Likelihood

ResearchDGX agent

arXiv:2605.29908v1 Announce Type: cross Abstract: Sparse recovery in linear systems underpins applications from signal processing to high-dimensional regression. Sparse Bayesian Learning, grounded in

Jony Ive’s funky Ferrari

IndustryDGX agent

Most people will never own, drive, or even sit inside a Ferrari Luce. (If you can, or do… hit us up.) There's still no question that Ferrari's first electric vehicle is one of the most interesting, su

Just dropped 🧑‍🍳 Fixed version of DeepSeek-V4-Pro-NVFP4 by @NVIDIAAI https://huggingface.co/nvidia/DeepSeek-V4-Pro-NVFP4

Model ReleasesDGX agent

NVIDIA has released a corrected version of DeepSeek-V4-Pro-NVFP4, a quantized model variant optimized for NVIDIA hardware using NV-FP4 (NVIDIA's 4-bit floating-point format). The model is available on

K-FinHallu: A Hallucination Detection Benchmark for Multi-Turn RAG in Korean Finance

Model ReleasesDGX agent

arXiv:2605.29523v1 Announce Type: new Abstract: Large Language Models (LLMs) have advanced financial automation through Retrieval-Augmented Generation (RAG), yet hallucinations remain a critical barri

KairosAgent: Agentic Time Series Forecasting with Fused Semantic Reasoning

AgentsDGX agent

arXiv:2605.30002v1 Announce Type: new Abstract: Cross-domain multimodal time series forecasting is a challenging task, requiring models to integrate precise numerical comprehension, cross-domain seman

Kalshi plans to offer perpetual futures contracts, saying it will be 'the first company in American history to offer' them, to be 'fully regulated' by the CFTC (Nathan Bomey/Axios)

IndustryDGX agent

Nathan Bomey / Axios: Kalshi plans to offer perpetual futures contracts, saying it will be “the first company in American history to offer” them, to be “fully regulated” by the CFTC — Kalshi announced

KAN-AD: Time Series Anomaly Detection with Kolmogorov-Arnold Networks

Local AiDGX agent

arXiv:2411.00278v4 Announce Type: replace Abstract: Time series anomaly detection (TSAD) underpins real-time monitoring in cloud services and web systems, allowing rapid identification of anomalies to

KBF: Knowledge Boundary as Fingerprint for Language Model and Black-Box API Auditing

Model ReleasesDGX agent

arXiv:2605.29524v1 Announce Type: cross Abstract: Relay and reseller APIs increasingly intermediate access to large language models (LLMs), but users have no direct way to verify that a claimed endpoi

Kernel-based potential mean-field games with unbiased random Fourier U-statistics

Model ReleasesDGX agent

arXiv:2605.29371v1 Announce Type: cross Abstract: We study the subclass of potential mean-field games in which the running interaction cost and the terminal target cost are both expressed through repr

Kernel Renormalization in Bayesian Deep Neural Networks: the Equivalent Wishart Ansatz in the Proportional Regime

Model ReleasesDGX agent

arXiv:2605.29684v1 Announce Type: new Abstract: The scaling limit where both the size of the training set P and the width N of a deep neural network grow at the same rate, the so-called proportional-w

KGEdit: Ambiguity-Aware Knowledge Graphs for Training-Free Precise Video Generation and Editing

SafetyDGX agent

arXiv:2605.29509v1 Announce Type: new Abstract: In recent years, training-free video generation has progressed remarkably. However, when handling complex textual instructions, existing methods still s

KLAS: Using Similarity to Stitch Neural Networks for Improved Accuracy-Efficiency Tradeoffs

ResearchDGX agent

arXiv:2605.29259v1 Announce Type: cross Abstract: Given the wide range of deployment targets, flexible model selection is essential for optimizing performance within a given compute budget. Recent wor

Knowing What to Solve Before How: Preplan Empowered LLM Mathematical Reasoning

TutorialsDGX agent

arXiv:2605.30245v1 Announce Type: new Abstract: Current plan-based reasoning methods improve large language models (LLMs) by inserting a planning stage before execution, giving rise to the question ri

Knowledge Offloading: Decomposing LLMs into Sparse Backbones and Memory Modules

Model ReleasesDGX agent

arXiv:2605.29075v1 Announce Type: new Abstract: LLMs encode both general capabilities and domain-specific knowledge in a single set of parameters. We ask whether this capacity can be reorganized: keep

Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models

Model ReleasesDGX agent

arXiv:2605.29459v1 Announce Type: new Abstract: Large language models route every input through a learned embedding table of shape |V| x d_model, consuming hundreds of millions to billions of trainabl

Label-Free Reinforcement Learning via Cross-Model Entropy

Model ReleasesDGX agent

arXiv:2605.29009v1 Announce Type: cross Abstract: Post-training large language models with reinforcement learning is bottlenecked by the reward signal. Existing approaches require either ground-truth

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs

Model ReleasesDGX agent

arXiv:2605.29928v1 Announce Type: cross Abstract: As AI-generated and AI-assisted content floods online spaces, source labels attached to such content can distort human reasoning judgments, with downs

LangSmith LLM Gateway lets you enforce spend limits and redacts PII before requests reach the model. Not after the fact.

AgentsDGX agent

LangSmith LLM Gateway is a feature that enables proactive cost and privacy controls by enforcing spending limits and redacting personally identifiable information (PII) before API requests are sent to

LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training

ResearchDGX agent

arXiv:2605.29888v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training has shown to improve reasoning in large language models (LLMs). However, there has been little exploration o

Large Depth Completion Model from Sparse Observations

TutorialsDGX agent

arXiv:2605.30115v1 Announce Type: new Abstract: This work presents the Large Depth Completion Model (LDCM), a simple, effective, and robust framework for single-view metric depth estimation with spars

Large language models reorganize representational geometry during in-context learning

Model ReleasesDGX agent

arXiv:2605.28854v1 Announce Type: new Abstract: Large language models (LLMs) exhibit remarkable flexibility: they can adapt to novel tasks from in-context examples without any parameter updates, a cap

Large-Scale AI and Foundation Models for Neuroscience: A Comprehensive Review

ResearchDGX agent

arXiv:2510.16658v3 Announce Type: replace Abstract: The development of large-scale artificial intelligence (AI) models is influencing neuroscience research by enabling end-to-end learning from raw bra

← Previous
1…784785786787788…1506
Next →