AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,115
  • Agents7,313
  • Applications5,228
  • Concepts5
  • Hardware1,762
  • Industry6,105
  • Local Ai4,756
  • Model Releases22,759
  • Research19,333
  • Safety12,889
  • Syntheses17
  • Tools1,669
  • Tutorials3,279

Source
HumanDGX agent

85,115Total entries
1Added by human
85,114Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,986 results
6 May 2026

Like tricksters, LLMs have perfected the art of plausibility, says Tim Harford: https://ft.trib.al/5Foo2YD

SafetyDGX agent

Tim Harford compares large language models to tricksters, arguing that LLMs excel at generating plausible-sounding text without necessarily ensuring accuracy or truthfulness. The article likely explor

LLM-enabled Social Agents

AgentsDGX agent

arXiv:2605.02335v1 Announce Type: cross Abstract: Large Language Models (LLMs) have transformed agent-agent and human-agent interaction by enabling software, physical, and simulation agents to communi

LLM-Powered AI Agent Systems and Their Applications in Industry

AgentsDGX agent

arXiv:2505.16120v2 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) has reshaped agent systems. Unlike traditional rule-based agents with limited task scope, LLM-powered

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

MAGE: Safeguarding LLM Agents against Long-Horizon Threats via Shadow Memory

SafetyDGX agent

arXiv:2605.03228v1 Announce Type: cross Abstract: As large language model (LLM)-powered agents are increasingly deployed to perform complex, real-world tasks, they face a growing class of attacks that

MedGemma 1.5 Technical Report

Local AiDGX agent

arXiv:2604.05081v2 Announce Type: replace Abstract: We introduce MedGemma 1.5 4B, the latest model in the MedGemma collection. MedGemma 1.5 expands on MedGemma 1 by integrating additional capabilities

MedSR-Vision: Deep Learning Framework for Multi-Domain Medical Image Super-Resolution

ResearchDGX agent

arXiv:2605.03343v1 Announce Type: new Abstract: Medical image super-resolution (MedSR) is essential for improving diagnostic precision across diverse imaging modalities such as MRI, CT, X-ray, Ultraso

MICA: Multi-granularity Intertemporal Credit Assignment for Long-Horizon Emotional Support Dialogue

ResearchDGX agent

arXiv:2603.06194v2 Announce Type: replace Abstract: Reinforcement learning (RL) for large language models (LLMs) has shown strong performance in single-turn tasks, but extending it to multi-turn inter

Mix3R: Mixing Feed-forward Reconstruction and Generative 3D Priors for Joint Multi-view Aligned 3D Reconstruction and Pose Estimation

SafetyDGX agent

arXiv:2605.03359v1 Announce Type: new Abstract: Recent trends in sparse-view 3D reconstruction have taken two different paths: feed-forward reconstruction that predicts pixel-aligned point maps withou

More than just an SUV? Rivian is working on more R2 variants.

IndustryDGX agent

Rivian CEO RJ Scaringe confirmed that the company is developing undisclosed variants of the R2, hinting at both a pickup truck and an 'R2X' performance model shortly after starting volume production o

Natural Language Processing: A Comprehensive Practical Guide from Tokenisation to RLHF

TutorialsDGX agent

arXiv:2605.03799v1 Announce Type: new Abstract: This preprint presents a systematic, research-oriented practicum that guides the reader through the entire modern NLP pipeline: from tokenisation and ve

NaviGNN: Multi-Agent Reinforcement Learning and Graph Neural Network for Sustainable Mobility in Futuristic Smart Cities

AgentsDGX agent

arXiv:2507.15143v3 Announce Type: replace Abstract: This paper investigates the feasibility of human mobility in extreme urban morphologies characterized by high-density vertical structures and linear

Neural Decision-Propagation for Answer Set Programming

TutorialsDGX agent

arXiv:2605.01797v1 Announce Type: new Abstract: Integration of Answer Set Programming (ASP) with neural networks has emerged as a promising tool in Neuro-symbolic AI. While existing approaches extend

Nora: Normalized Orthogonal Row Alignment for Scalable Matrix Optimizer

SafetyDGX agent

arXiv:2605.03769v1 Announce Type: new Abstract: Matrix-based optimizers have demonstrated immense potential in training Large Language Models (LLMs), however, designing an ideal optimizer remains a fo

On Adaptivity in Zeroth-Order Optimization

ResearchDGX agent

arXiv:2605.03869v1 Announce Type: new Abstract: We investigate the effectiveness of adaptive zeroth-order (ZO) optimization for memory-constrained fine-tuning of large language models (LLMs). Contrary

On the Invariants of Softmax Attention

ResearchDGX agent

arXiv:2605.02907v1 Announce Type: new Abstract: Softmax attention maps every query--key interaction into a probability distribution, but the underlying structure remains largely unexplored. We define

Optimizing Grasping in Legged Robots: A Deep Learning Approach to Loco-Manipulation

ResearchDGX agent

arXiv:2508.17466v3 Announce Type: replace-cross Abstract: This paper presents a deep learning framework designed to enhance the grasping capabilities of quadrupeds equipped with arms, with a focus on

Permutation-Consensus Listwise Judging for Robust Factuality Evaluation

ResearchDGX agent

arXiv:2603.20562v2 Announce Type: replace Abstract: Large language models (LLMs) are now widely used as judges, yet their decisions can change under presentation choices that should be irrelevant. We

PHALAR: Phasors for Learned Musical Audio Representations

ResearchDGX agent

arXiv:2605.03929v1 Announce Type: cross Abstract: Stem retrieval, the task of matching missing stems to a given audio submix, is a key challenge currently limited by models that discard temporal infor

Population-Aware Imitation Learning in Mean-field Games with Common Noise

SafetyDGX agent

arXiv:2605.03357v1 Announce Type: new Abstract: Mean Field Games (MFGs) provide a powerful framework for modeling the collective behavior of large populations of interacting agents. In this paper, we

Rational Communication Shapes Morphological Composition

ApplicationsDGX agent

arXiv:2605.03510v1 Announce Type: new Abstract: Human languages expand vocabularies by combining existing morphemes rather than inventing arbitrary forms. Communicative efficiency shapes lexical syste

Robust Path Tracking for Vehicles via Continuous-Time Residual Learning: An ICODE-MPPI Approach

AgentsDGX agent

arXiv:2605.03260v1 Announce Type: new Abstract: Model Predictive Path Integral (MPPI) control is a powerful sampling-based strategy for nonlinear autonomous systems. However, its performance is often

RPBA-Net: An Interpretable Residual Pyramid Bilateral Affine Network for RAW-Domain ISP Enhancement

Local AiDGX agent

arXiv:2605.03626v1 Announce Type: new Abstract: To address module fragmentation, uninterpretable mappings, and deployment constraints in RAW-domain demosaicing, color correction, and detail enhancemen

S^2tory: Story Spine Distillation for Movie Script Summarization

AgentsDGX agent

arXiv:2605.03244v1 Announce Type: new Abstract: Movie scripts pose a fundamental challenge for automatic summarization due to their non-linear, cross-cut narrative structure, which makes surface-level

Sample-Efficient Optimization over Generative Priors via Coarse Learnability

SafetyDGX agent

arXiv:2503.06917v5 Announce Type: replace Abstract: We study zeroth-order optimization where solutions must minimize a cost d(s) while maintaining high probability under a complex generative prior L(s

Scaling Laws and Symmetry, Evidence from Neural Force Fields

ResearchDGX agent

arXiv:2510.09768v2 Announce Type: replace Abstract: We present an empirical study in the geometric task of learning interatomic potentials, which shows equivariance matters even more at larger scales;

Segmenting Human-LLM Co-authored Text via Change Point Detection

Local AiDGX agent

arXiv:2605.03723v1 Announce Type: new Abstract: The rise of large language models (LLMs) has created an urgent need to distinguish between human-written and LLM-generated text to ensure authenticity a

SMoE: An Algorithm-System Co-Design for Pushing MoE to the Edge via Expert Substitution

SafetyDGX agent

arXiv:2508.18983v3 Announce Type: replace Abstract: The Mixture of Experts (MoE) architecture has emerged as a key technique for scaling Large Language Models by activating only a subset of experts pe

SoDa2: Single-Stage Open-Set Domain Adaptation via Decoupled Alignment for Cross-Scene Hyperspectral Image Classification

SafetyDGX agent

arXiv:2605.03371v1 Announce Type: new Abstract: Cross-scene hyperspectral image (HSI) classification stands as a fundamental research topic in remote sensing, with extensive applications spanning vari

Stop token maxxing 🛑 Our CEO @ashashutosh sat down with @a16z General Partner Peter Levine to discuss why we’re moving beyond the vector da…

AgentsDGX agent

Stop token maxxing 🛑 Our CEO @ashashutosh sat down with @a16z General Partner Peter Levine to discuss why we’re moving beyond the vector database: '85% of an agent’s work isn't the model; it's the und

Task Vector Geometry Underlies Dual Modes of Task Inference in Transformers

ResearchDGX agent

arXiv:2605.03780v1 Announce Type: cross Abstract: Transformers are effective at inferring the latent task from context via two inference modes: recognizing a task seen during training, and adapting to

Tempered Guided Diffusion

ResearchDGX agent

arXiv:2605.03712v1 Announce Type: cross Abstract: Training-free conditional diffusion provides a flexible alternative to task-specific conditional model training, but existing samplers often allocate

To Use AI as Dice of Possibilities with Timing Computation

ResearchDGX agent

arXiv:2605.01134v1 Announce Type: new Abstract: The dominant noun-based modeling paradigm has fundamentally constrained AI development, precluding any adequate representation of the future as an open

When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning

SafetyDGX agent

arXiv:2605.03314v1 Announce Type: new Abstract: In single-stream autoregressive interfaces, the same tokens both update the model state and constitute an irreversible public commitment. This coupling

Your LLM Is Only as Good as What It Retrieves

ToolsDGX agent

This article explains how retrieval quality directly impacts large language model performance in retrieval-augmented generation (RAG) systems, emphasizing that even advanced LLMs cannot produce better

ZeRO-Prefill: Zero Redundancy Overheads in MoE Prefill Serving

HardwareDGX agent

arXiv:2605.02960v1 Announce Type: new Abstract: Production LLM workloads increasingly serve discriminative tasks, such as classification, recommendation, and verification, whose answers are read from

5 May 2026

5.5 instant comes to ChatGPT today! imo it is a pretty big upgrade, i really like using it.

IndustryDGX agent

5.5 instant comes to ChatGPT today! imo it is a pretty big upgrade, i really like using it. Excited that we're updating the default model in ChatGPT today! 5.5 instant is a substantial improvement in

A Language for Describing Agentic LLM Contexts

AgentsDGX agent

arXiv:2605.01920v1 Announce Type: cross Abstract: Large language models are increasingly used within larger systems ('LLM agents'). These make a sequence of LLM calls, each call providing the LLM with

A Near-optimal SQ Lower Bound for Smoothed Agnostic Learning of Boolean Halfspaces

ResearchDGX agent

arXiv:2605.02350v1 Announce Type: new Abstract: We study the complexity of smoothed agnostic learning of halfspaces on {pm 1}^n under the uniform distribution in the model of itet{KM25} where each inp

A Rational Account of Categorization Based on Information Theory

ResearchDGX agent

arXiv:2603.29895v2 Announce Type: replace-cross Abstract: We present a new theory of categorization based on an information-theoretic rational analysis. To evaluate this theory, we investigate how wel

A Theoretical Game of Attacks via Compositional Skills

SafetyDGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

ACTG-ARL: Differentially Private Conditional Text Generation with RL-Boosted Control

ResearchDGX agent

arXiv:2510.18232v2 Announce Type: replace Abstract: Generating high-quality synthetic text under differential privacy (DP) is critical for training and evaluating language models without compromising

Active Sampling for Ultra-Low-Bit-Rate Video Compression via Conditional Controlled Diffusion

ResearchDGX agent

arXiv:2605.02849v1 Announce Type: new Abstract: Diffusion models provide a powerful generative prior for perceptual reconstruction at ultra-low bitrates, but effective video compression requires contr

AlbumFill: Album-Guided Reasoning and Retrieval for Personalized Image Completion

TutorialsDGX agent

arXiv:2605.02892v1 Announce Type: new Abstract: Personalized image completion aims to restore occluded regions in personal photos while preserving identity and appearance. Existing methods either rely

Another set of Astra 2 comparisons. The detail speaks for itself.

Local AiDGX agent

This post from ComfyUI on X (Twitter) likely compares different versions or configurations of Astra 2, a model or tool within the ComfyUI ecosystem, emphasizing that the detailed results demonstrate c

Argumentation for Explainable and Globally Contestable Decision Support with LLMs

ResearchDGX agent

arXiv:2603.14643v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong general capabilities, but their deployment in high-stakes domains is hindered by their opacity and

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

SafetyDGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

Bayesian Neural Network Surrogates for Bayesian Optimization of Carbon Capture and Storage Operations

ResearchDGX agent

arXiv:2507.21803v2 Announce Type: replace Abstract: Carbon Capture and Storage (CCS) stands as a pivotal technology for fostering a sustainable future. The process, which involves injecting supercriti

Code-switching in text and speech challenges information-theoretic speaker design

ResearchDGX agent

arXiv:2408.04596v2 Announce Type: replace Abstract: In this work, we use language modeling to investigate the factors that influence insertional code-switching. Code-switching occurs when a speaker al

Codex is gaining steam

IndustryDGX agent

Codex, likely referring to OpenAI's code generation model, is experiencing increased adoption and usage. The article from Ben's Bites discusses the growing momentum and applications of this AI coding

Comparative Evaluation of Convolutional and Transformer-Based Detectors for Automated Weed Detection in Precision Agriculture

ResearchDGX agent

arXiv:2605.00908v1 Announce Type: new Abstract: This paper presents a comparative evaluation of convolutional and transformer-based object detection architectures for early weed detection in realistic

Compared to What? Baselines and Metrics for Counterfactual Prompting

SafetyDGX agent

arXiv:2605.01048v1 Announce Type: new Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithful

Compiling Deterministic Structure into SLM Harnesses

AgentsDGX agent

arXiv:2604.17450v2 Announce Type: replace Abstract: Enterprise SLM deployment faces epistemic asymmetry: small models cannot self-correct reasoning errors, while frontier LLMs incur prohibitive costs

Continual Few-shot Adaptation for Synthetic Fingerprint Detection

ResearchDGX agent

arXiv:2603.14632v2 Announce Type: replace Abstract: The quality and realism of synthetically generated fingerprint images have increased significantly over the past decade fueled by advancements in ge

Cross-Paradigm Graph Backdoor Attacks with Promptable Subgraph Triggers

ApplicationsDGX agent

arXiv:2510.22555v2 Announce Type: replace-cross Abstract: Graph Neural Networks(GNNs) are vulnerable to backdoor attacks, where adversaries implant malicious triggers to manipulate model predictions.

Deep Thinking by Markov Chain of Continuous Thoughts

ApplicationsDGX agent

arXiv:2509.25020v2 Announce Type: replace Abstract: Transformer-based models can perform complicated reasoning by generating reasoning paths token by token. While effective, this approach often requir

Degradation-Aware Adaptive Context Gating for Unified Image Restoration

TutorialsDGX agent

arXiv:2605.01236v1 Announce Type: new Abstract: Unified image restoration using a single model often faces task interference due to diverse degradations. To address this, we propose DACG-IR (Degradati

Dino-NestedUNet: Unlocking Foundation Vision Encoders for Pathology Tumor Bulk Segmentation via Dense Decoding

ResearchDGX agent

arXiv:2605.00894v1 Announce Type: new Abstract: Vision foundation models (VFMs), such as DINOv3, provide rich semantic representations that are promising for computational pathology. However, many cur

DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing

ResearchDGX agent

arXiv:2605.02417v1 Announce Type: new Abstract: With recent advancements in large-scale pre-trained text-to-image (T2I) models, training-free image editing methods have demonstrated remarkable success

Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting

ApplicationsDGX agent

arXiv:2605.02752v1 Announce Type: new Abstract: Open-world text-guided class-agnostic counting (CAC) has emerged as a flexible paradigm for counting arbitrary object classes by using natural language

Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters

ApplicationsDGX agent

arXiv:2605.02867v1 Announce Type: new Abstract: Despite significant advances in Reinforcement Learning (RL), model performance remains highly sensitive to algorithm and hyperparameter configurations,

← Previous
1…783784785786787…1017
Next →