AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,060
  • Agents7,763
  • Applications5,542
  • Concepts5
  • Hardware1,932
  • Industry6,210
  • Local Ai5,103
  • Model Releases24,798
  • Research20,784
  • Safety13,745
  • Syntheses17
  • Tools1,680
  • Tutorials3,481

Source
HumanDGX agent

Content type
91,060Total entries
1Added by human
91,059Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,793 results
Agents

GenesisFunc: Multi-Agent Data Generation for Accurate and Generalizable Function-Calling

DGX agent

arXiv:2605.28835v1 Announce Type: cross Abstract: Large Language Models (LLMs) extend their capabilities through function-calling (FC), which relies on training data with high quality, diversity, and

agentsarxiv-cs-ai
29 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

GiPL: Generative augmented iterative Pseudo-Labeling for Cross-Domain Few-Shot Object Detection

DGX agent

arXiv:2605.29539v1 Announce Type: cross Abstract: Vision-language foundation models have shown promising zero-shot generalization for Cross-Domain Few-Shot Object Detection (CD-FSOD). However, they fa

researcharxiv-cs-ai
29 May 2026
Model Releases

Give it Space! Explicit Disentangling of Positional and Semantic Representations in Encoders

DGX agent

arXiv:2605.30022v1 Announce Type: cross Abstract: Positional encoding (PE) underpins how permutation-invariant Transformers represent sequence order, yet how positional information is processed and st

model-releasesarxiv-cs-ai
29 May 2026
Research

Gradient Preconditioning for Efficient and Reliable Reward-Guided Generation

DGX agent

arXiv:2602.08646v2 Announce Type: replace Abstract: We propose a gradient preconditioning method that makes reward-guided generation with one-step generative models both efficient and reliable. Test-t

researcharxiv-cs-lg
29 May 2026
Model Releases

GroundAct: Can LLM Agents Ground Actions in Environmental States?

DGX agent

arXiv:2508.05614v2 Announce Type: replace-cross Abstract: LLM agents achieve 85-96% success on tasks where instructions fully specify the action, but drop to 29-53% when action feasibility depends on

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing

DGX agent

arXiv:2605.29532v1 Announce Type: cross Abstract: Exploratory GUI testing is a particularly demanding setting for MLLM agents: without predefined test scripts, an agent must autonomously navigate an a

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Hallucination Mitigation with Agentic AI, Nested Learning, and AI Sustainability via Semantic Caching

DGX agent

arXiv:2605.29055v1 Announce Type: new Abstract: Hallucination remains a major reliability barrier for production LLM systems, particularly in multi-agent pipelines where unsupported claims can propaga

model-releasesarxiv-cs-ai
29 May 2026
Research

HaluNet: Learning Hallucination Risk from Internal Signals in LLM Question Answering

DGX agent

arXiv:2512.24562v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong question answering (QA) performance but can produce fluent answers unsupported by available evidence. Ex

researcharxiv-cs-cl
29 May 2026
Model Releases

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, …

DGX agent

Hear the architects of Gemini reflect on their journey to continue pushing the frontier of AI, on this episode of Release Notes. @JeffDean, @koraykv, @OriolVinyalsML, and @NoamShazeer sit down on came

model-releasesgoogle-ai--x
29 May 2026
Model Releases

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that th…

DGX agent

I’m going to let you in on a secret. I made a mistake once.* In August 2024.* It’s true. I actually got something wrong. I predicted that there would be a collapse of the AI bubble (which in my judgem

model-releasesgary-marcus--x
29 May 2026
Model Releases

Improving Adversarial Robustness of Attribution via Implicit Regularization

DGX agent

arXiv:2605.29983v1 Announce Type: cross Abstract: The adversarial robustness of attributions is a fundamental requirement for reliable explainability in deep learning, yet existing approaches typicall

model-releasesarxiv-cs-cv
29 May 2026
Research

In-Place Feedback: Reliable Refinement for Multi-Turn Expert-LLM Collaboration

DGX agent

arXiv:2510.00777v2 Announce Type: replace Abstract: LLM-generated drafts often contain subtle factual or logical errors, yet prior work shows that models struggle to reliably integrate multi-turn feed

researcharxiv-cs-lg
29 May 2026
Model Releases

InsightEval: An Expert-Curated Benchmark for Assessing Insight Discovery in LLM-Driven Data Agents

DGX agent

arXiv:2511.22884v2 Announce Type: replace Abstract: Data analysis has become an indispensable part of scientific research. To discover the latent knowledge and insights hidden within massive datasets,

model-releasesarxiv-cs-ai
29 May 2026
Agents

KairosAgent: Agentic Time Series Forecasting with Fused Semantic Reasoning

DGX agent

arXiv:2605.30002v1 Announce Type: new Abstract: Cross-domain multimodal time series forecasting is a challenging task, requiring models to integrate precise numerical comprehension, cross-domain seman

agentsarxiv-cs-ai
29 May 2026
Model Releases

Label Over Logic? How Source Cues Bias Human Fallacy Judgments More Than LLMs

DGX agent

arXiv:2605.29928v1 Announce Type: cross Abstract: As AI-generated and AI-assisted content floods online spaces, source labels attached to such content can distort human reasoning judgments, with downs

model-releasesarxiv-cs-ai
29 May 2026
Research

LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training

DGX agent

arXiv:2605.29888v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training has shown to improve reasoning in large language models (LLMs). However, there has been little exploration o

researcharxiv-cs-ai
29 May 2026
Tutorials

Latent Terms: Dense Retrievers Contain Trivially Extractable BM25-ready Zipfian Vocabularies

DGX agent

arXiv:2605.29384v1 Announce Type: cross Abstract: We propose Latent Terms, a method revealing that models trained for dense retrieval, whether single- or multi-vector, learn representations that can t

tutorialsarxiv-cs-ai
29 May 2026
Model Releases

Learning Representations from 3D Gaussian Splats

DGX agent

arXiv:2605.29549v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) is a recent approach for scene rendering. Although primarily designed for view synthesis, its potential for scene understan

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

LiteCoder-Terminal: Scaling Long-Horizon Terminal Environments for Learning Language Agents

DGX agent

arXiv:2605.29559v1 Announce Type: new Abstract: Mastering terminal environments requires language agents capable of multi-step planning, feedback-grounded execution, and dynamic state adaptation. Howe

model-releasesarxiv-cs-cl
29 May 2026
Model Releases

LLM-Evolved Domain-Independent Heuristics for Symbolic AI Planning

DGX agent

arXiv:2605.29649v1 Announce Type: new Abstract: Heuristic search is the dominant paradigm in symbolic AI planning, and the strongest heuristics are the result of decades of work by planning researcher

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LoCoT2V-Bench: Benchmarking Long-Form and Complex Text-to-Video Generation

DGX agent

arXiv:2510.26412v3 Announce Type: replace-cross Abstract: Recent advances in text-to-video generation have achieved impressive performance on short clips, yet evaluating long-form generation under com

model-releasesarxiv-cs-ai
29 May 2026
Research

MIRAGE: Adaptive Multimodal Gating for Whole-Brain fMRI Encoding

DGX agent

arXiv:2605.29850v1 Announce Type: new Abstract: Recent progress in task-optimized neural networks has established encoding models as a powerful tool for predicting brain responses to naturalistic stim

researcharxiv-cs-lg
29 May 2026
Model Releases

MPDocBench-Parse: Benchmarking Practical Multi-page Document Parsing

DGX agent

arXiv:2605.22100v2 Announce Type: replace Abstract: Document parsing converts visually rich documents into machine-readable structured representations, forming a crucial foundation for information sys

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization

DGX agent

arXiv:2605.29951v1 Announce Type: new Abstract: Understanding how harm emerges from interaction between otherwise benign image-text pairs requires intent-aware cross-modal reasoning beyond surface-lev

model-releasesarxiv-cs-ai
29 May 2026
Research

One Click per Cell Type Suffices: Training-free Group Interaction for Cell Instance Segmentation

DGX agent

arXiv:2605.29429v1 Announce Type: new Abstract: Cell instance segmentation models trained on cell-specific datasets suffer severe performance drops on out-of-distribution cell types, while interactive

researcharxiv-cs-cv
29 May 2026
Model Releases

OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

DGX agent

arXiv:2605.29253v1 Announce Type: new Abstract: Task success can hide process anomalies in real-world agent executions. An agent may pass the final task oracle while still accumulating unresolved ambi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

OVA-IB: One vs All Information Bottleneck for Multi-Modal Alignment

DGX agent

arXiv:2605.29900v1 Announce Type: new Abstract: Contrastive learning is effective for aligning paired views or modalities, but alignment beyond two modalities remains non-trivial and comparatively und

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Parameter-Efficient Subspace Decoupling ViT for Mitigating Multi-Task Negative Transfer in Histological Scoring

DGX agent

arXiv:2605.29852v1 Announce Type: new Abstract: Histological scoring is essential for diagnosing Non-Alcoholic Fatty Liver Disease (NAFLD), yet its automation remains challenging due to the high annot

model-releasesarxiv-cs-cv
29 May 2026
Applications

Persona Conditioning of Brand Recommendations in Retrieval-Augmented Commercial Chat: A Prominence-Stratified Cross-Provider Audit

DGX agent

arXiv:2605.30207v1 Announce Type: new Abstract: The same prompt -- 'best CRM software' -- reaches AI assistants from buyers in widely different contexts: a solo founder, an enterprise VP, a UK SMB own

applicationsarxiv-cs-ai
29 May 2026
Model Releases

PhAIL: A Real-Robot VLA Benchmark and Distributional Methodology

DGX agent

arXiv:2605.29710v1 Announce Type: new Abstract: Real-world evaluation of vision-language-action (VLA) policies still rests on binary success rate at a fixed timeout with N le 25 rollouts per condition

model-releasesarxiv-cs-ro
29 May 2026
Model Releases

PokerSkill: LLMs Can Play Expert-Level Poker without Training or Solvers

DGX agent

arXiv:2605.30094v1 Announce Type: new Abstract: Poker is a landmark challenge for artificial intelligence. The dominant approach relies on equilibrium solvers built on counterfactual regret minimizati

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Pre-Registering the Detectable Effect: A Paired-MDE Budget for 4-bit Quantization Benchmarks, with a Pilot Audit

DGX agent

arXiv:2605.28873v1 Announce Type: new Abstract: This is a planning-method note with an unpaired pilot audit. We adapt the classical paired-binary sample-size calculation (Miettinen, 1968) to quantizat

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

Prescribe-then-Select: Adaptive Policy Selection for Contextual Stochastic Optimization

DGX agent

arXiv:2509.08194v2 Announce Type: replace Abstract: We address the problem of policy selection in contextual stochastic optimization (CSO), where covariates are available as contextual information and

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

PTCG-Bench: Can LLM Agents Master Pokemon Trading Card Game?

DGX agent

arXiv:2605.29653v1 Announce Type: new Abstract: Given a strategically complex board game, human players can quickly learn to devise strategies after playing a few rounds. Autonomous agents require sim

model-releasesarxiv-cs-ai
29 May 2026
Safety

Recurrent Structural Policy Gradient for Partially Observable Mean Field Games

DGX agent

arXiv:2602.20141v2 Announce Type: replace Abstract: Mean Field Games (MFGs) provide a principled framework for modelling interactions in large population systems. However, algorithmic progress has bee

safetyarxiv-cs-ai
29 May 2026
Safety

Reinforcing Few-step Generators via Reward-Tilted Distribution Matching

DGX agent

arXiv:2605.26108v2 Announce Type: replace Abstract: Recent advances in few-step diffusion distillation have enabled efficient image generation, yet aligning these models with human preferences remains

safetyarxiv-cs-cv
29 May 2026
Model Releases

SAAS: Self-Aware Reinforcement Learning for Over-Search Mitigation in Agentic Search

DGX agent

arXiv:2605.29796v1 Announce Type: new Abstract: Agentic search enables LLMs to solve complex multi-hop questions through iterative reasoning and external search. Despite the effectiveness, these syste

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

SAGE: Segment-Aware Gloss-Free Encoding for Token-Efficient Sign Language Translation

DGX agent

arXiv:2507.09266v2 Announce Type: replace Abstract: Gloss-free Sign Language Translation (SLT) has advanced rapidly, achieving strong performances without relying on gloss annotations. However, these

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

Sample-Efficient Diffusion-based Reinforcement Learning with Critic Guidance

DGX agent

arXiv:2605.30056v1 Announce Type: cross Abstract: Recent advances in reinforcement learning (RL) have achieved great successes by leveraging the multimodality and exploration capability of diffusion p

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

SDF-Net: Structure-Aware Disentangled Feature Learning for Opticall-SAR Ship Re-identification

DGX agent

arXiv:2603.12588v2 Announce Type: replace Abstract: Cross-modal ship re-identification (ReID) between optical and synthetic aperture radar (SAR) imagery is fundamentally challenged by the severe radio

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

STAMP: Training Explicit Memory for Mobile GUI Agents in Controllable and Scalable Virtual Environments

DGX agent

arXiv:2605.29324v1 Announce Type: new Abstract: Mobile GUI agents excel at immediate reactive control but frequently fail in realistic, long-horizon tasks that require memory. This failure stems from

model-releasesarxiv-cs-cl
29 May 2026
Safety

Statistical Embeddings for Similarity, Retrieval, and Interpretable Alignment of Numeric Tabular Datasets

DGX agent

arXiv:2605.30289v1 Announce Type: new Abstract: Numeric tabular datasets are the dominant data format in scientific practice, yet large language models lack native mechanisms for representing numeric

safetyarxiv-cs-lg
29 May 2026
Research

Streaming Drag-Oriented Interactive Video Manipulation: Drag Anything, Anytime!

DGX agent

arXiv:2510.03550v4 Announce Type: replace Abstract: Achieving streaming, fine-grained control over the outputs of autoregressive video diffusion models remains challenging, making it difficult to ensu

researcharxiv-cs-cv
29 May 2026
Model Releases

Striding Across Reynolds Numbers: Representation Geometry in Neural PDE Generalisation

DGX agent

arXiv:2605.30112v1 Announce Type: new Abstract: Cross-Reynolds generalisation in neural PDE solvers remains poorly characterised. On the canonical forced 2D Navier-Stokes benchmark, a trained Fourier

model-releasesarxiv-cs-lg
29 May 2026
Model Releases

TAE: Target-aware enhancer for nighttime UAV tracking

DGX agent

arXiv:2605.29558v1 Announce Type: new Abstract: Severe image degradation under low-light nighttime conditions constitutes a core bottleneck preventing all-day applications for UAV-based single object

model-releasesarxiv-cs-cv
29 May 2026
Model Releases

TANDEM: Temporal-Aware Neural Detection for Multimodal Hate Speech

DGX agent

arXiv:2601.11178v2 Announce Type: replace Abstract: Social media platforms are increasingly dominated by long-form multimodal content, where harmful narratives are constructed through a complex interp

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

Text-Preserving Lossy Text Compression: A Study of Strategic Deletion and LLM Reconstruction

DGX agent

arXiv:2605.29000v1 Announce Type: new Abstract: Traditional lossless text compression preserves every byte, but its gains on natural language are often modest in realistic operating regimes. We study

model-releasesarxiv-cs-cl
29 May 2026
Research

The Little Book of Generative AI Foundations: An Intuitive Mathematical Primer

DGX agent

arXiv:2605.29713v1 Announce Type: cross Abstract: This book provides a compact, derivation-oriented introduction to the mathematical foundations of modern generative artificial intelligence. Rather th

researcharxiv-cs-ai
29 May 2026
← Previous
1…707708709710711…1371
Next →