AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,016 results
9 Jun 2026

GimmBO: Interactive Generative Image Model Merging via Bayesian Optimization

ApplicationsDGX agent

arXiv:2601.18585v2 Announce Type: replace Abstract: Fine-tuning-based adaptation is widely used to customize diffusion-based image generation, leading to large collections of community-created adapter

Graph-GRPO: Training Graph Flow Models with Reinforcement Learning

ResearchDGX agent

arXiv:2603.10395v2 Announce Type: replace Abstract: Graph generation is a fundamental task with broad applications, such as drug discovery. Recently, discrete flow matching-based graph generation, aka

HARP: Efficient Data Selection for Finetuning Large Language Models

SafetyDGX agent

arXiv:2606.07690v1 Announce Type: cross Abstract: Finetuning data selection requires balancing two competing goals: selecting examples that improve the downstream objective, and doing so without repea

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Large Language Models for Imbalanced Classification: Diversity makes the difference

ResearchDGX agent

arXiv:2510.09783v2 Announce Type: replace-cross Abstract: Oversampling is one of the most widely used approaches for addressing imbalanced classification. The core idea is to generate additional minor

Learning to Attack and Defend: Adaptive Red Teaming of Language Models via GRPO

SafetyDGX agent

arXiv:2606.09701v1 Announce Type: cross Abstract: AI red teaming must continually adapt to evolving attackers and defenders. Reinforcement learning offers a promising approach to discovering novel att

Optimality of Sequential Filtering Under Independent Cost and Selectivity Models

ResearchDGX agent

arXiv:2606.07589v1 Announce Type: new Abstract: Sequential filtering pipelines are a common design pattern in large-scale systems, where a large population of items is progressively reduced by a seque

Reason Twice: Segmentation via Candidate Discovery and Comparative Reasoning

Model ReleasesDGX agent

arXiv:2606.09303v1 Announce Type: new Abstract: The rapid development of pretrained foundation models has enabled more general image segmentation. Multimodal large language models (MLLMs) have been wi

Simultaneous hyperkinetic movement disorders phenotyping: a cross-cohort pediatric transfer study using routine videos, markerless pose estimation and a tabular foundation model

ApplicationsDGX agent

arXiv:2606.07674v1 Announce Type: new Abstract: Objective: To develop and externally test a video-based framework for simultaneous detection of hyperkinetic MDs phenomenologies: dystonia, tremor, myoc

Single-Cell Cross-Modal Transfer by Adversarial Fine-Tuning of Foundation Models

ResearchDGX agent

arXiv:2606.07676v1 Announce Type: cross Abstract: Spatial transcriptomics (ST) is a powerful tool for exploring biological properties dependent on structure, proximity, and interaction in tissue. The

When Vision Misleads, Let Location Speak: A Worldwide Image Geo-Localization Method via Location Attention Mechanism and Large Multimodal Models

Local AiDGX agent

arXiv:2606.08918v1 Announce Type: new Abstract: Worldwide image geo-localization aims to determine the capture location of an image on a global scale. Existing methods often mislocalize images by matc

Where Does the Answer Come From? Benchmarking View-Level Visual Evidence Identification in Multi-View MLLMs for Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.09644v1 Announce Type: cross Abstract: Multimodal large language models (MLLMs) achieve strong results on visual reasoning benchmarks, but answer accuracy alone does not indicate whether a

8 Jun 2026

Auditing Training Data in Domain-adapted LLMs: LoRA-MINT

Model ReleasesDGX agent

arXiv:2606.06946v1 Announce Type: cross Abstract: We present LoRA-MINT, a new methodology for Membership Inference Test (MINT) applied to recent Large Language Models (LLMs) fine-tuned for specific Na

dots.tts Technical Report

Model ReleasesDGX agent

arXiv:2606.07080v1 Announce Type: cross Abstract: We present dots.tts, a 2B-parameter continuous autoregressive text-to-speech (TTS) foundation model that models speech in a continuous latent space. C

Hearing the Unspoken: Language Model Priors for Acoustic Adversarial Attacks

ResearchDGX agent

arXiv:2606.06833v1 Announce Type: cross Abstract: Automatic Speech Recognition (ASR) systems operating in real-time settings must process acoustic input under strict temporal constraints, where transc

TorchKM: A GPU-Oriented Library for Kernel Learning and Model Selection

HardwareDGX agent

arXiv:2606.06742v1 Announce Type: new Abstract: TorchKM is an open-source library for kernel machines, including support vector machines, kernel logistic regression, and kernel quantile regression, wi

Trace Reconstruction with Language Models

ApplicationsDGX agent

arXiv:2507.12927v2 Announce Type: replace Abstract: The general trace reconstruction problem seeks to recover an original sequence from its noisy copies independently corrupted by insertions, deletion

6 Jun 2026

Critic-Guided Heterogeneous Multi-Agent Reasoning for Reliable Mathematical Problem Solving

Model ReleasesDGX agent

arXiv:2606.05704v1 Announce Type: new Abstract: Recent Large Language Models (LLMs) have shown impressive reasoning abilities; but they are still susceptible to hallucinations, intermediate reasoning

Exploring LLMs for South Asian Music Understanding and Generation

Model ReleasesDGX agent

arXiv:2606.05522v1 Announce Type: cross Abstract: Recent advancements in Large Language Models (LLMs) have shown promising results in music understanding and generation tasks. However, existing works

Safety Paradox: How Enhanced Safety Awareness Leaves LLMs Vulnerable to Posterior Attack

Model ReleasesDGX agent

arXiv:2606.05614v1 Announce Type: new Abstract: Large language models (LLMs) are rigorously aligned to refuse harmful requests, a process that inherently cultivates a latent capacity to evaluate and r

5 Jun 2026

A Conversational Framework for Human-Robot Collaborative Manipulation with Distributed Generative AI models

Local AiDGX agent

arXiv:2606.06061v1 Announce Type: new Abstract: This paper presents a distributed conversational framework for human-robot collaborative manipulation that integrates local language and vision-language

Correcting Prompt Dependence in LLM Benchmarks: A Bayesian Hierarchical Model with Embedding-Space Clustering

ResearchDGX agent

arXiv:2510.05709v2 Announce Type: replace-cross Abstract: LLM benchmarking metrics often misstate performance and uncertainty as they rely on two assumptions that frequently do not hold in practice: (

Deep Learning-assisted AMD Staging based on OCT and OCT Angiography

ResearchDGX agent

arXiv:2606.05379v1 Announce Type: new Abstract: To develop and evaluate deep learning models for automated grading of age-related macular degeneration (AMD) severity using optical coherence tomography

Multi-Task Crack Foundation Model for Engineering-Reliable Crack Representation and Topology Preservation in Civil Infrastructure

ResearchDGX agent

arXiv:2606.05641v1 Announce Type: new Abstract: Reliable crack assessment requires not only accurate pixel-level masks but also connected crack geometry and confidence estimates that remain stable und

4 Jun 2026

FoeGlass: Simple In-Context Learning Is Enough for Red Teaming Audio Deepfake Detectors

ResearchDGX agent

arXiv:2606.05101v1 Announce Type: cross Abstract: Audio deepfake detection (ADD) models are critical for countering the malicious use of text-to-speech (TTS) models. Evaluating and strengthening ADD m

Learning symplectic model reduction based on a approximation theorem of symplectic embeddings

ResearchDGX agent

arXiv:2606.04623v1 Announce Type: new Abstract: High-dimensional Hamiltonian systems play a central role in many scientific and engineering disciplines, with dynamics evolving on symplectic manifolds.

Long Live Fine-Tuning: Task-Specific Transformers Outperform Zero-Shot LLMs for Misinformation Response Classification on Reddit

Model ReleasesDGX agent

arXiv:2606.04274v1 Announce Type: new Abstract: As large language models (LLMs) become default tools for online information verification, an implicit assumption follows them: that scale and general ca

Not All Errors Are Equal: Consequence-Aware Reasoning Compute Allocation

Model ReleasesDGX agent

arXiv:2606.04402v1 Announce Type: new Abstract: Modern reasoning models can allocate different amounts of test-time computation, such as thinking tokens, model calls, or compute budget, to different t

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is th…

Model ReleasesDGX agent

Outstanding paper on long-horizon agents. (bookmark it) Similar to humans, how do you make agents persist on a difficult task, and how is that useful? And which models today work well on this? This ne

Physics-Informed Machine Learning for Short-Term Flood Prediction

SafetyDGX agent

arXiv:2606.04143v1 Announce Type: cross Abstract: Accurate flood forecasting is essential for mitigating disaster risks and protecting communities. However, purely data-driven machine learning models

Self-Evaluation Is Already There: Eliciting Latent Judge Calibration in Base LLMs with Minimal Data

Local AiDGX agent

arXiv:2606.05122v1 Announce Type: new Abstract: Large language models are increasingly evaluated by other models, raising a natural question: can a model predict how a judge will score its own output?

Toward a Generalized Defense Across Sparse, Continuous, and Structured Parameter Attacks

Model ReleasesDGX agent

arXiv:2606.04317v1 Announce Type: cross Abstract: Deep neural networks are increasingly deployed across heterogeneous and partially untrusted environments, where models are distributed through cloud s

Trusted healthcare AI hinges on data foundations, not models alone

ApplicationsDGX agent

Healthcare AI is moving out of the pilot phase and into production environments, where the gap between a compelling demo and a clinically trustworthy output has never been more consequential. As AI ag

WAM-Nav: Asymmetric Latent World-Action Modeling for Unified Visual Navigation

SafetyDGX agent

arXiv:2606.04907v1 Announce Type: new Abstract: Visual navigation requires generating smooth and collision-free trajectories under complex geometric and physical constraints. Existing reactive policie

3 Jun 2026

A Scoping Review of the Ethical Perspectives on Anthropomorphising Large Language Model-Based Conversational Agents

ResearchDGX agent

arXiv:2601.09869v2 Announce Type: replace Abstract: Anthropomorphisation -- the phenomenon whereby non-human entities are ascribed human-like qualities -- has become increasingly salient with the rise

AirDreamer: Generalist Drone Navigation with World Models

SafetyDGX agent

arXiv:2606.03252v1 Announce Type: cross Abstract: Navigating a drone in unseen and cluttered environments requires reliable generalization to unseen scene layouts and understanding of environmental st

APIC: Amortized Physics-Informed Calibration using Neural Processes

Model ReleasesDGX agent

arXiv:2606.03355v1 Announce Type: new Abstract: Physics models are inherently imperfect due to misspecified or missing mechanisms, resulting in systematic discrepancies between model predictions and r

Are we really tilting? The mechanics of reward guidance in flow and diffusion models

SafetyDGX agent

arXiv:2606.02884v1 Announce Type: cross Abstract: Reward guidance algorithms steer a learned generative process toward the reward-tilted measure at inference time. While empirically powerful, these me

Balancing Symmetry and Efficiency in Graph Flow Matching

ResearchDGX agent

arXiv:2602.18084v2 Announce Type: replace Abstract: Equivariance is central to graph generative models, as it ensures the model respects the permutation symmetry of graphs. However, strict equivarianc

Benchmarking Speech-to-Speech Translation Models

ApplicationsDGX agent

arXiv:2606.03241v1 Announce Type: new Abstract: Speech-to-speech translation (S2ST) has advanced rapidly, but offline evaluation lacks a unified protocol: studies report non-overlapping metric subsets

COD10K-C: Benchmarking Robustness of Camouflaged Object Detection Under Natural Image Corruptions

Model ReleasesDGX agent

arXiv:2606.02603v1 Announce Type: new Abstract: Camouflaged object detection has improved substantially, but most standard benchmarks evaluate models only on clean images. This is not realistic becaus

Diagnosing Knowledge Gaps in LLM Tool Use: An Agentic Benchmark for Novel API Acquisition

Model ReleasesDGX agent

arXiv:2606.03657v1 Announce Type: new Abstract: Large language models for code generation often need to use APIs that are absent from their pretraining data. This requires more than recalling a functi

Disentangling Visual and Factual Correctness in LVLMs' Visualization Literacy

Model ReleasesDGX agent

arXiv:2606.03142v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) show strong visualization interpretation, yet it is unclear whether their responses reflect genuine reasoning over

E2LLM: Towards Efficient LLM Serving in Heterogeneous Edge/Fog Environments

ApplicationsDGX agent

arXiv:2606.03770v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become integral to modern applications, yet their deployment remains challenging. Beyond executing the models themse

Enginuity: A Dataset and Benchmark for Vision-Language Understanding of Engineering Diagrams

Model ReleasesDGX agent

arXiv:2606.03410v1 Announce Type: new Abstract: Engineering diagrams pose a distinct challenge for vision-language models: unlike natural images or general documents, they encode information through d

FinStressTS: A Parametric Synthetic Benchmark for Time-Series Forecasting in Finance

Model ReleasesDGX agent

arXiv:2606.03184v1 Announce Type: cross Abstract: Financial forecasting is difficult due to low signal-to-noise ratios, latent factors, heavy tails, regime shifts, and jumps. Real-world benchmarks off

Fixed-Time Dynamic Landing of Quadrotors using Adaptive Unscented Kalman Filtering and Nonlinear Model Predictive Control

ResearchDGX agent

arXiv:2606.02658v1 Announce Type: new Abstract: This paper introduces an estimation and control framework for dynamic landing of multi-rotor uncrewed aerial vehicles on moving platforms. The proposed

Follow-Your-Preference++: Rethinking Preference Alignment for Image Inpainting

SafetyDGX agent

arXiv:2606.03216v1 Announce Type: new Abstract: We study preference alignment for image inpainting. Rather than proposing yet another method, we revisit the problem from first principles and reassess

Gender-Dependent Diagnostic Substitution in LLM Medical Triage: Same Symptoms, Unequal Urgency

Model ReleasesDGX agent

arXiv:2606.03641v1 Announce Type: new Abstract: We investigate whether large language models produce different medical triage recommendations for identical neurological symptoms when only the patient'

GeoDrive-Bench: Benchmarking Region-Specific Multimodal Reasoning in Autonomous Driving

Model ReleasesDGX agent

arXiv:2606.02774v1 Announce Type: new Abstract: Vision-language models (VLMs) for autonomous driving have shown promising performance, but their ability to handle region-specific traffic rules remains

Inverting the Generation Process of Denoising Diffusion Implicit Models: Empirical Evaluation and a Novel Method

ResearchDGX agent

arXiv:2606.03111v1 Announce Type: new Abstract: This paper studies the problem of inverting the DDIM image generation process to recover latent variables, particularly the initial noise map, from a ge

Laplacian Representations for Decision-Time Planning

Model ReleasesDGX agent

arXiv:2602.05031v2 Announce Type: replace Abstract: Planning with a learned model remains a key challenge in model-based reinforcement learning (RL). In decision-time planning, state representations a

Rethinking Molecular Text Representations for LLMs: An Empirical Study

Model ReleasesDGX agent

arXiv:2606.03057v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used for molecular tasks, but it remains unclear which molecular representation to use. We present a sys

Skill-RM: Unifying Heterogeneous Evaluation Criteria via Agent Skill

Model ReleasesDGX agent

arXiv:2606.03980v1 Announce Type: cross Abstract: Reward models (RMs) provide critical feedback signals for LLM post-training, notably in reinforced fine-tuning (RFT) and reinforcement learning (RL) p

Trump plan to test AI models has a problem—US security teams were gutted by DOGE

IndustryDGX agent

Trump signed an executive order on June 2, 2026, establishing a voluntary framework for the federal government to evaluate risks posed by advanced AI systems before release and encouraging companies t

Video-Mirai: Autoregressive Video Diffusion Models Need Foresight

TutorialsDGX agent

arXiv:2606.03971v1 Announce Type: new Abstract: Causal video generators must predict from the past, but they need not learn only from it. In streaming autoregressive video diffusion, each emitted segm

2 Jun 2026

A Protocol-Language Model for Network Intrusion (Without Deep Packet Inspection)

SafetyDGX agent

arXiv:2606.00155v1 Announce Type: cross Abstract: Modern network intrusion detection systems (NIDS) are caught in a structural contradiction: the protocols carrying the highest threat intelligence are

A Structured Benchmark for Text-Guided Anomaly Detection: When Language Stops Conditioning the Decision

Model ReleasesDGX agent

arXiv:2606.01992v1 Announce Type: cross Abstract: Industrial anomaly detection has historically been a unimodal task. Recent multimodal vision-language models have produced systems that admit textual

BraveGuard: From Open-World Threats to Safer Computer-Use Agents

Model ReleasesDGX agent

arXiv:2606.01166v1 Announce Type: cross Abstract: Computer-use agents extend language models from text generation to sustained interaction with files, terminals, browsers, and external tools. This shi

ChWDTA: Channel-wise Wavelet-Domain Transformer Attention and Entropy Modeling for Learned Image Compression

ResearchDGX agent

arXiv:2606.00111v1 Announce Type: cross Abstract: State-of-the-art learned image compression (LIC) schemes are increasingly based on hybrid CNN-transformer architectures. To further improve rate-disto

Correcting Gradient-Based Circuit Localization via Interaction-Aware Backpropagation

Model ReleasesDGX agent

arXiv:2505.17630v4 Announce Type: replace Abstract: Circuit localization methods aim to identify the subset of model components responsible for specific behaviors in large language models, enabling de

← Previous
1…256257258259260…1034
Next →