AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Research

When Tables Leak: Attacking String Memorization in LLM-Based Tabular Data Generation

DGX agent

arXiv:2512.08875v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have recently demonstrated remarkable performance in generating high-quality tabular synthetic data. In practice,

researcharxiv-cs-ai
12 May 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Why Retrying Fails: Context Contamination in LLM Agent Pipelines

DGX agent

arXiv:2605.08563v1 Announce Type: new Abstract: When an LLM agent fails a multi-step tool-augmented task and retries, the failed attempt typically remains in its context window -- contaminating the ne

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

2.5-D Decomposition for LLM-Based Spatial Construction

DGX agent

arXiv:2605.07066v1 Announce Type: new Abstract: Autonomous systems that build structures from natural-language instructions need reliable spatial reasoning, yet large language models (LLMs) make syste

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Ask Patients with Patience: Enabling LLMs for Human-Centric Medical Dialogue with Grounded Reasoning

DGX agent

arXiv:2502.07143v3 Announce Type: replace Abstract: The severe shortage of medical doctors limits access to timely and reliable healthcare, leaving millions underserved. Large language models (LLMs) o

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

Beyond Factor Aggregation: Gauge-Aware Low-Rank Server Representations for Federated LoRA

DGX agent

arXiv:2605.06733v1 Announce Type: cross Abstract: Federated LoRA enables parameter-efficient adaptation of large language models under decentralized data and limited client resources.However, directly

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Beyond Factual Accuracy: Evaluating Global Reasoning Integrity in RAG Systems with LogicScore

DGX agent

arXiv:2601.15050v4 Announce Type: replace Abstract: Current evaluation methods for Retrieval Augmented Generation (RAG) suffer from extit{factual myopia}: they relentlessly emphasize factual accuracy

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

BRIDGE: Background Routing and Isolated Discrete Gating for Coarse-Mask Local Editing

DGX agent

arXiv:2605.07846v1 Announce Type: new Abstract: Coarse-mask local image editing asks a model to modify a user-indicated region while preserving the surrounding scene. In practice, however, rough masks

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning

DGX agent

arXiv:2605.07251v1 Announce Type: new Abstract: Large Language Models (LLMs) have become increasingly capable as tool-using agents, with benchmarks spanning diverse general agentic tasks. Yet rigorous

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ChartREG++: Towards Benchmarking and Improving Chart Referring Expression Grounding under Diverse referring clues and Multi-Target Referring

DGX agent

arXiv:2605.07415v1 Announce Type: cross Abstract: Referring expression grounding is a core problem in visual grounding and is widely used as a diagnostic of spatial grounding and reasoning in vision a

model-releasesarxiv-cs-cl
11 May 2026
Applications

Christoffel-DPS: Optimal sensor placement in diffusion posterior sampling for arbitrary distributions

DGX agent

arXiv:2605.06861v1 Announce Type: new Abstract: State estimation is a critical task in scientific, engineering and control applications. Since the reliability of reconstructions depends on the number

applicationsarxiv-cs-lg
11 May 2026
Model Releases

Contrast-X: A Multi-Modal Contrast Image Synthesis Benchmark and Universal Modality Flow Matching

DGX agent

arXiv:2601.15884v2 Announce Type: replace Abstract: Contrast-enhanced imaging is central to oncologic diagnosis, but contrast agents can be contraindicated for many of the patients who need them most.

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

CSMCIR: CoT-Enhanced Symmetric Alignment with Memory Bank for Composed Image Retrieval

DGX agent

arXiv:2601.03728v3 Announce Type: replace-cross Abstract: Composed Image Retrieval (CIR) enables users to search for target images using both a reference image and manipulation text, offering substant

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

CyBiasBench: Benchmarking Bias in LLM Agents for Cyber-Attack Scenarios

DGX agent

arXiv:2605.07830v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in offensive cybersecurity. In this paper, we reveal an interesting phenom

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Excluding the Target Domain Improves Extrapolation: Deconfounded Hierarchical Physics Constraints

DGX agent

arXiv:2605.07485v1 Announce Type: cross Abstract: Extrapolation to out-of-distribution conditions is a fundamental challenge for physics-constrained deep generative models. Existing methods apply phys

model-releasesarxiv-cs-ai
11 May 2026
Tutorials

ForgeVLA: Federated Vision-Language-Action Learning without Language Annotations

DGX agent

arXiv:2605.07474v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models hold great promise for general-purpose robotic intelligence, yet scaling up such models is severely bottlenecked b

tutorialsarxiv-cs-ai
11 May 2026
Model Releases

From Synthetic to Real: Toward Identity-Consistent Makeup Transfer with Synthetic and Real Data

DGX agent

arXiv:2605.07861v1 Announce Type: new Abstract: Makeup transfer aims to apply the makeup style of a reference portrait to a source portrait while preserving identity and background. Early methods form

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

GAD in the Wild: Benchmarking Graph Anomaly Detection under Realistic Deployment Challenges

DGX agent

arXiv:2605.07133v1 Announce Type: cross Abstract: Graph Anomaly Detection (GAD) is a critical task in graph machine learning with vital applications in financial fraud detection and social platform go

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GazeVLM: Active Vision via Internal Attention Control for Multimodal Reasoning

DGX agent

arXiv:2605.07817v1 Announce Type: cross Abstract: Human visual reasoning is governed by active vision, a process where metacognitive control drives top-down goal-directed attention, dynamically routin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

GraphReAct: Reasoning and Acting for Multi-step Graph Inference

DGX agent

arXiv:2605.07357v1 Announce Type: new Abstract: Reasoning-acting frameworks enhance large language models (LLMs) by interleaving reasoning with actions for dynamic information acquisition. However, ex

model-releasesarxiv-cs-ai
11 May 2026
Research

Identifiability Challenges in Sparse Linear Ordinary Differential Equations

DGX agent

arXiv:2506.09816v3 Announce Type: replace Abstract: Dynamical systems modeling is a core pillar of scientific inquiry across natural and life sciences. Increasingly, dynamical system models are learne

researcharxiv-cs-lg
11 May 2026
Local Ai

Is She Even Relevant? When BERT Ignores Explicit Gender Cues

DGX agent

arXiv:2605.07622v1 Announce Type: new Abstract: Gender bias in large language models has primarily been investigated for English, while languages with grammatical or morphological gender remain compar

local-aiarxiv-cs-cl
11 May 2026
Model Releases

LARAG: Link-Aware Retrieval Strategy for RAG Systems in Hyperlinked Technical Documentation

DGX agent

arXiv:2605.07517v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) enhances the factual grounding of Large Language Models by conditioning their outputs on external documents. Howe

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

MatryoshkaLoRA: Learning Accurate Hierarchical Low-Rank Representations for LLM Fine-Tuning

DGX agent

arXiv:2605.07850v1 Announce Type: cross Abstract: With the rise in scale for deep learning models to billions of parameters, the computational cost of fine-tuning remains a significant barrier to depl

model-releasesarxiv-cs-ai
11 May 2026
Applications

Multimodal synthesis of MRI and tabular data with diffusion in a joint latent space via cross-attention

DGX agent

arXiv:2605.06699v1 Announce Type: cross Abstract: We propose a multimodal latent diffusion model that jointly synthesizes volumetric magnetic resonance imaging (MRI) and tabular clinical data within a

applicationsarxiv-cs-ai
11 May 2026
Model Releases

MultiSoc-4D: A Benchmark for Diagnosing Instruction-Induced Label Collapse in Closed-Set LLM Annotation of Bengali Social Media

DGX agent

arXiv:2605.06940v1 Announce Type: new Abstract: Annotation automation via Large Language Models (LLMs) is the core approach for scaling NLP datasets; however, LLM behavior with respect to closed-set i

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

NCL-UoR at SemEval-2026 Task 5: Embedding-Based Methods, Fine-Tuning, and LLMs for Word Sense Plausibility Rating

DGX agent

arXiv:2603.08256v2 Announce Type: replace Abstract: Word sense plausibility rating requires predicting the human-perceived plausibility of a given word sense on a 1-5 scale in the context of short nar

model-releasesarxiv-cs-cl
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Safety

On Training in Imagination

DGX agent

arXiv:2605.06732v1 Announce Type: new Abstract: State-of-the-art model-based reinforcement learning methods train policies on imagined rollouts. These rollouts are trajectories generated by a learned

safetyarxiv-cs-lg
11 May 2026
Model Releases

OpenAI launches professional services business with $4B investment

DGX agent

OpenAI Group PBC today unveiled a new business unit, The OpenAI Deployment Company, that will help companies adopt its artificial intelligence models. The subsidiary is launching with 4 billion in fun

model-releasessiliconangle
11 May 2026
Model Releases

PerCaM-Health: Personalized Dynamic Causal Graphs for Healthcare Reasoning

DGX agent

arXiv:2605.07267v1 Announce Type: new Abstract: Personalized healthcare decisions require reasoning about how physiological and behavioral variables influence an individual patient over time. Existing

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

ProcObject-10K: Benchmarking Object-Centric Procedural Understanding in Instructional Videos

DGX agent

arXiv:2512.03479v2 Announce Type: replace Abstract: Procedural activities are fundamentally driven by object state transitions, yet existing instructional video benchmarks remain action-centric and ca

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

Prompt Engineering Strategies for LLM-based Qualitative Coding of Psychological Safety in Software Engineering Communities: A Controlled Empirical Study

DGX agent

arXiv:2605.07422v1 Announce Type: cross Abstract: Qualitative analysis plays a pivotal role in understanding the human and social aspects of software engineering. However, it remains a demanding proce

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to …

DGX agent

Qwen 3.6 Plus by @Alibaba_Qwen is now FREE for a limited time on Nous Portal! Nous Portal is one easy subscription that gives you access to 300+ models, exclusive discounts, and bundles your tokens an

model-releasesnous-research--x
11 May 2026
Safety

ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning

DGX agent

arXiv:2605.07477v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have achieved remarkable progress, however, many edited results still suffer from artifacts, unintended mo

safetyarxiv-cs-cv
11 May 2026
Agents

RelAgent: LLM Agents as Data Scientists for Relational Learning

DGX agent

arXiv:2605.07840v1 Announce Type: new Abstract: Relational learning is a challenging problem that has motivated a wide range of approaches, including graph-based models (e.g., graph neural networks, g

agentsarxiv-cs-lg
11 May 2026
Model Releases

ReSeek: A Self-Correcting Framework for Search Agents with Instructive Rewards

DGX agent

arXiv:2510.00568v3 Announce Type: replace Abstract: Search agents powered by Large Language Models (LLMs) have demonstrated significant potential in tackling knowledge-intensive tasks. Reinforcement l

model-releasesarxiv-cs-cl
11 May 2026
Model Releases

RRCM: Ranking-Driven Retrieval over Collaborative and Meta Memories for LLM Recommendation

DGX agent

arXiv:2605.07129v1 Announce Type: cross Abstract: Large Language Models (LLMs) have emerged as a promising paradigm for next-generation recommender systems, offering strong semantic understanding and

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

RuleSafe-VL: Evaluating Rule-Conditioned Decision Reasoning in Vision-Language Content Moderation

DGX agent

arXiv:2605.07760v1 Announce Type: new Abstract: Platform content moderation applies explicit policy rules and context-dependent conditions to decide whether user content is allowed, restricted, or rem

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Scaling Continual Learning to 300+ Tasks with Bi-Level Routing Mixture-of-Experts

DGX agent

arXiv:2602.03473v2 Announce Type: replace-cross Abstract: Continual learning, especially class-incremental learning (CIL), on the basis of a pre-trained model (PTM) has garnered substantial research i

model-releasesarxiv-cs-cv
11 May 2026
Model Releases

SCOPE: Structured Decomposition and Conditional Skill Orchestration for Complex Image Generation

DGX agent

arXiv:2605.08043v1 Announce Type: cross Abstract: While text-to-image models have made strong progress in visual fidelity, faithfully realizing complex visual intents remains challenging because many

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Self-Play Enhancement via Advantage-Weighted Refinement in Online Federated LLM Fine-Tuning with Real-Time Feedback

DGX agent

arXiv:2605.07977v1 Announce Type: new Abstract: Recent works have advanced feedback-based learning systems, whereby a foundation model is able to intake incoming feedback (e.g., a user) to self-improv

model-releasesarxiv-cs-lg
11 May 2026
Research

SHRED: Retain-Set-Free Unlearning via Self-Distillation with Logit Demotion

DGX agent

arXiv:2605.07482v1 Announce Type: cross Abstract: Machine unlearning for large language models (LLMs) aims to selectively remove memorized content such as private data, copyrighted text, or hazardous

researcharxiv-cs-ai
11 May 2026
Research

Stochastic Transition-Map Distillation for Fast Probabilistic Inference

DGX agent

arXiv:2605.07661v1 Announce Type: cross Abstract: Diffusion models achieve strong generation quality, diversity, and distribution coverage, but their performance often comes with expensive inference.

researcharxiv-cs-cv
11 May 2026
Model Releases

Structure Over Scale: Learning Visual Reasoning from Pedagogical Video

DGX agent

arXiv:2601.23251v2 Announce Type: replace Abstract: State-of-the-art vision-language models (VLMs) score impressively on video benchmarks yet stumble on basic visual reasoning tasks involving spatial

model-releasesarxiv-cs-cv
11 May 2026
Local Ai

Teaching Prompts to Coordinate: Hierarchical Layer-Grouped Prompt Tuning for Continual Learning

DGX agent

arXiv:2511.12090v3 Announce Type: replace Abstract: Prompt-based continual learning methods fine-tune only a small set of additional learnable parameters while keeping the pre-trained model's paramete

local-aiarxiv-cs-cv
11 May 2026
Model Releases

Text-to-CAD Evaluation with CADTests

DGX agent

arXiv:2605.07807v1 Announce Type: cross Abstract: Text-to-CAD has recently emerged as an important task with the potential to substantially accelerate design workflows. Despite its significance, there

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

The Translation Tax Is Not a Scalar: A Counterfactual Audit of English-Source Cue Inheritance in Chinese Multilingual Benchmarks

DGX agent

arXiv:2605.07093v1 Announce Type: cross Abstract: The Translation Tax is often treated as a scalar: translated benchmarks are assumed to inflate scores by preserving English-source cues. We audit this

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…542543544545546…1371
Next →