AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries90,316
  • Agents7,714
  • Applications5,508
  • Concepts5
  • Hardware1,905
  • Industry6,191
  • Local Ai5,052
  • Model Releases24,539
  • Research20,616
  • Safety13,635
  • Syntheses17
  • Tools1,678
  • Tutorials3,456

Source
HumanDGX agent

Content type
AllBlog
90,316Total entries
1Added by human
90,315Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,172 results
Model Releases

TabReX : Tabular Referenceless eXplainable Evaluation

DGX agent

arXiv:2512.15907v2 Announce Type: replace Abstract: Evaluating the quality of tables generated by large language models (LLMs) remains an open challenge: existing metrics either flatten tables into te

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Take Out Your Calculators: Estimating the Real Difficulty of Question Items with LLM Student Simulations

DGX agent

arXiv:2601.09953v2 Announce Type: replace Abstract: Standardized math assessments require expensive human pilot studies to establish the difficulty of test items. We investigate the predictive value o

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Talking to a Know-It-All GPT or a Second-Guesser Claude? How Repair reveals unreliable Multi-Turn Behavior in LLMs

DGX agent

arXiv:2604.19245v1 Announce Type: cross Abstract: Repair, an important resource for resolving trouble in human-human conversation, remains underexplored in human-LLM interaction. In this study, we inv

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

The new Gemini Enterprise: one platform for agent development, orchestration, and governance

DGX agent

The first wave of AI changed how we find information; the next wave is changing how we get work done. Today, we’re enhancing our most powerful AI tools and bringing them together under one roof. Gemin

model-releasesgoogle-cloud-ai
22 Apr 2026
Research

TrEEStealer: Stealing Decision Trees via Enclave Side Channels

DGX agent

arXiv:2604.18716v1 Announce Type: cross Abstract: Today, machine learning is widely applied in sensitive, security-related, and financially lucrative applications. Model extraction attacks undermine c

researcharxiv-cs-lg
22 Apr 2026
Research

Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length

DGX agent

arXiv:2603.22608v2 Announce Type: replace Abstract: Users often rely on Large Language Models (LLMs) for processing multiple documents or performing analysis over a number of instances. For example, a

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Unveiling Fine-Grained Visual Traces: Evaluating Multimodal Interleaved Reasoning Chains in Multimodal STEM Tasks

DGX agent

arXiv:2604.19697v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have shown promising reasoning abilities, yet evaluating their performance in specialized domains remains chall

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

VDPP: Video Depth Post-Processing for Speed and Scalability

DGX agent

arXiv:2604.06665v2 Announce Type: replace Abstract: Video depth estimation is essential for providing 3D scene structure in applications ranging from autonomous driving to mixed reality. Current end-t

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

DGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

What’s new in the Agentic Data Cloud: Powering the System of Action

DGX agent

Companies are shifting from gen AI that simply answers questions to autonomous agents that perceive, reason, and act on their behalf. Attempting to scale these agents on legacy stacks exposes structur

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

When and What to Ask: AskBench and Rubric-Guided RLVR for LLM Clarification

DGX agent

arXiv:2602.11199v2 Announce Type: replace Abstract: Large language models (LLMs) often respond even when prompts omit critical details or include misleading information, leading to hallucinations or r

model-releasesarxiv-cs-cl
22 Apr 2026
Agents

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool mode…

DGX agent

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool model with `hermes tools` Introducing ChatGPT Images 2.0 A state

agentsnous-research--x
21 Apr 2026
Model Releases

Adaptive Local Frequency Filtering for Fourier-Encoded Implicit Neural Representations

DGX agent

arXiv:2604.02846v2 Announce Type: replace Abstract: Fourier-encoded implicit neural representations (INRs) have shown strong capability in modeling continuous signals from discrete samples. However, c

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning

DGX agent

arXiv:2604.17889v1 Announce Type: new Abstract: Despite recent progress in multimodal large language models (MLLMs), reliable visual question answering in aerial scenes remains challenging. In such sc

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AeroScene: Progressive Scene Synthesis for Aerial Robotics

DGX agent

arXiv:2603.23224v2 Announce Type: replace Abstract: Generative models have shown substantial impact across multiple domains, their potential for scene synthesis remains underexplored in robotics. This

model-releasesarxiv-cs-ro
21 Apr 2026
Model Releases

ArgBench: Benchmarking LLMs on Computational Argumentation Tasks

DGX agent

arXiv:2604.17366v1 Announce Type: new Abstract: Argumentation skills are an essential toolkit for large language models (LLMs). These skills are crucial in various use cases, including self-reflection

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Benchmarking programs?

DGX agent

The Reddit post 'Benchmarking programs?' in r/ollama likely discusses tools and methods for measuring the performance of local language models running on Ollama. Available benchmarking tools for Ollam

local-air-ollama
21 Apr 2026
Local Ai

Beyond Memorization: Extending Reasoning Depth with Recurrence, Memory and Test-Time Compute Scaling

DGX agent

arXiv:2508.16745v2 Announce Type: replace Abstract: Reasoning is a core capability of large language models, yet how multi-step reasoning is learned and executed remains unclear. We study this questio

local-aiarxiv-cs-lg
21 Apr 2026
Research

Bridging Coarse and Fine Recognition: A Hybrid Approach for Open-Ended Multi-Granularity Object Recognition in Interactive Educational Games

DGX agent

arXiv:2604.16785v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have enabled open-ended object recognition, yet they struggle with fine-grained tasks. In co

researcharxiv-cs-cv
21 Apr 2026
Applications

Calibrated? Not for Everyone: How Sexual Orientation and Religious Markers Distort LLM Accuracy and Confidence in Medical QA

DGX agent

arXiv:2604.17316v1 Announce Type: new Abstract: Safe clinical deployment of Large Language Models (LLMs) requires not only high accuracy but also robust uncertainty calibration to ensure models defer

applicationsarxiv-cs-cl
21 Apr 2026
Applications

Can we generate portable representations for clinical time series data using LLMs?

DGX agent

arXiv:2603.23987v2 Announce Type: replace Abstract: Deploying clinical ML is slow and brittle: models that work at one hospital often degrade under distribution shifts at the next. In this work, we st

applicationsarxiv-cs-lg
21 Apr 2026
Model Releases

CaseFacts: A Benchmark for Legal Fact-Checking and Precedent Retrieval

DGX agent

arXiv:2601.17230v2 Announce Type: replace Abstract: Automated Fact-Checking has largely focused on verifying general knowledge against static corpora, overlooking high-stakes domains like law where tr

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

CBRS: Cognitive Blood Request System with Bilingual Dataset and Dual-Layer Filtering for Multi-Platform Social Streams

DGX agent

arXiv:2604.16665v1 Announce Type: new Abstract: Urgent blood donation seeking posts and messages on social media often go unnoticed due to the overwhelming volume of daily communications. Traditional

model-releasesarxiv-cs-cl
21 Apr 2026
Research

Coevolving Representations in Joint Image-Feature Diffusion

DGX agent

arXiv:2604.17492v1 Announce Type: new Abstract: Joint image-feature generative modeling has recently emerged as an effective strategy for improving diffusion training by coupling low-level VAE latents

researcharxiv-cs-cv
21 Apr 2026
Agents

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

DGX agent

arXiv:2509.00789v2 Announce Type: replace Abstract: The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs):

agentsarxiv-cs-cv
21 Apr 2026
Research

Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding

DGX agent

arXiv:2511.08480v3 Announce Type: replace Abstract: Multimodal Large Language Models advance multimodal representation learning by acquiring transferable semantic embeddings, thereby substantially enh

researcharxiv-cs-cv
21 Apr 2026
Research

DisCa: Accelerating Video Diffusion Transformers with Distillation-Compatible Learnable Feature Caching

DGX agent

arXiv:2602.05449v3 Announce Type: replace Abstract: While diffusion models have achieved great success in the field of video generation, this progress is accompanied by a rapidly escalating computatio

researcharxiv-cs-cv
21 Apr 2026
Safety

DMax: Aggressive Parallel Decoding for dLLMs

DGX agent

arXiv:2604.08302v2 Announce Type: replace Abstract: We present DMax, a new paradigm for efficient diffusion language models (dLLMs). It mitigates error accumulation in parallel decoding, enabling aggr

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Do LLM-derived graph priors improve multi-agent coordination?

DGX agent

arXiv:2604.17191v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) is crucial for AI systems that operate collaboratively in distributed and adversarial settings, particularly i

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Document-as-Image Representations Fall Short for Scientific Retrieval

DGX agent

arXiv:2604.18508v1 Announce Type: cross Abstract: Many recent document embedding models are trained on document-as-image representations, embedding rendered pages as images rather than the underlying

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

Downgrade to Upgrade: Optimizer Simplification Enhances Robustness in LLM Unlearning

DGX agent

arXiv:2510.00761v5 Announce Type: replace Abstract: Large language model (LLM) unlearning aims to surgically remove the influence of undesired data or knowledge from an existing model while preserving

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

EgoSound: Benchmarking Sound Understanding in Egocentric Videos

DGX agent

arXiv:2602.14122v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have recently achieved remarkable progress in vision-language understanding. Yet, human perception is inher

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Enhancing Glass Surface Reconstruction via Depth Prior for Robot Navigation

DGX agent

arXiv:2604.18336v1 Announce Type: cross Abstract: Indoor robot navigation is often compromised by glass surfaces, which severely corrupt depth sensor measurements. While foundation models like Depth A

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

EVE: Verifiable Self-Evolution of MLLMs via Executable Visual Transformations

DGX agent

arXiv:2604.18320v1 Announce Type: new Abstract: Self-evolution of multimodal large language models (MLLMs) remains a critical challenge: pseudo-label-based methods suffer from progressive quality degr

safetyarxiv-cs-cv
21 Apr 2026
Model Releases

EvoCoT: Overcoming the Exploration Bottleneck in Reinforcement Learning

DGX agent

arXiv:2508.07809v5 Announce Type: replace Abstract: Reinforcement learning with verifiable reward (RLVR) has become a promising paradigm for post-training large language models (LLMs) to improve their

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

FireScope: Wildfire Risk Prediction with a Chain-of-Thought Oracle

DGX agent

arXiv:2511.17171v4 Announce Type: replace Abstract: Predicting wildfire risk is a reasoning-intensive spatial problem that requires the integration of visual, climatic, and geographic factors to infer

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

DGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

model-releasesopenai--x
21 Apr 2026
Research

Forget What Matters, Keep the Rest: Selective Unlearning of Informative Tokens

DGX agent

arXiv:2604.17785v1 Announce Type: new Abstract: Unlearning in large language models (LLMs) has emerged as a promising safeguard against adversarial behaviors. When the forgetting loss is applied unifo

researcharxiv-cs-cl
21 Apr 2026
Model Releases

FRIGID: Scaling Diffusion-Based Molecular Generation from Mass Spectra at Training and Inference Time

DGX agent

arXiv:2604.16648v1 Announce Type: new Abstract: In this work, we present FRIGID, a framework with a novel diffusion language model that generates molecular structures conditioned on mass spectra via i

model-releasesarxiv-cs-lg
21 Apr 2026
Research

From Attribution to Abstention: Training-Free Attention-Based Auditing for Clinical Summarization

DGX agent

arXiv:2601.16397v2 Announce Type: replace Abstract: Deploying multimodal large language models (MLLMs) for clinical summarization demands not only fluent generation but also transparency about where e

researcharxiv-cs-cl
21 Apr 2026
Model Releases

From log pi to pi: Taming Divergence in Soft Clipping via Bilateral Decoupled Decay of Probability Gradient Weight

DGX agent

arXiv:2603.14389v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has catalyzed a leap in Large Language Model (LLM) reasoning, yet its optimization dynamics re

model-releasesarxiv-cs-lg
21 Apr 2026
Safety

Geometric Stability: The Missing Axis of Representations

DGX agent

arXiv:2601.09173v4 Announce Type: replace-cross Abstract: Representational similarity analysis and related methods have become standard tools for comparing the internal geometries of neural networks a

safetyarxiv-cs-cl
21 Apr 2026
Safety

GeometryZero: Advancing Geometry Solving via Group Contrastive Policy Optimization

DGX agent

arXiv:2506.07160v3 Announce Type: replace Abstract: Recent progress in large language models (LLMs) has boosted mathematical reasoning, yet geometry remains challenging where auxiliary construction is

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

GR4CIL: Gap-compensated Routing for CLIP-based Class Incremental Learning

DGX agent

arXiv:2604.17822v1 Announce Type: new Abstract: Class-Incremental Learning (CIL) aims to continuously acquire new categories while preserving previously learned knowledge. Recently, Contrastive Langua

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

DGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

DGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

iDocV2: Leveraging Self-Supervision and Open-Set Detection for Improving Pattern Spotting in Historical Documents

DGX agent

arXiv:2604.16726v1 Announce Type: new Abstract: Considering the imminent massification of digital books, it has become critical to facilitate searching collections through graphical patterns. Current

model-releasesarxiv-cs-cv
21 Apr 2026
Research

iPhoneme: Brain-to-Text Communication for ALS Using ConformerXL Decoding

DGX agent

arXiv:2604.16441v1 Announce Type: cross Abstract: Brain-computer interfaces (BCIs) for speech restoration hold transformative potential for the approximately 173,000--232,500 individuals worldwide wit

researcharxiv-cs-cl
21 Apr 2026
← Previous
1…474475476477478…1358
Next →