AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,429
  • Agents7,559
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,170
  • Local Ai4,935
  • Model Releases23,918
  • Research20,125
  • Safety13,371
  • Syntheses17
  • Tools1,677
  • Tutorials3,403

Source
HumanDGX agent

Content type
AllBlog
88,429Total entries
1Added by human
88,428Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,648 results
Research

R-CoV: Region-Aware Chain-of-Verification for Alleviating Object Hallucinations in LVLMs

DGX agent

arXiv:2604.20696v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have demonstrated impressive performance in various multimodal understanding and reasoning tasks. However, they sti

researcharxiv-cs-cv
23 Apr 2026
Tutorials
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Render-in-the-Loop: Vector Graphics Generation via Visual Self-Feedback

DGX agent

arXiv:2604.20730v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have shown promising capabilities in generating Scalable Vector Graphics (SVG) via direct code synthesis. Howev

tutorialsarxiv-cs-cv
23 Apr 2026
Research

Semantic-Fast-SAM: Efficient Semantic Segmenter

DGX agent

arXiv:2604.20169v1 Announce Type: new Abstract: We propose Semantic-Fast-SAM (SFS), a semantic segmentation framework that combines the Fast Segment Anything model with a semantic labeling pipeline to

researcharxiv-cs-cv
23 Apr 2026
Model Releases

SpeechParaling-Bench: A Comprehensive Benchmark for Paralinguistic-Aware Speech Generation

DGX agent

arXiv:2604.20842v1 Announce Type: cross Abstract: Paralinguistic cues are essential for natural human-computer interaction, yet their evaluation in Large Audio-Language Models (LALMs) remains limited

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

SweRank: Software Issue Localization with Code Ranking

DGX agent

arXiv:2505.07849v2 Announce Type: replace-cross Abstract: Software issue localization, the task of identifying the precise code locations (files, classes, or functions) relevant to a natural language

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

Think fast!

DGX agent

Think fast! Introducing Grok Voice Think Fast 1.0 A state-of-the-art voice model built for complex, multi-step workflows with snappy responses and high accuracy. It takes the top spot on the Tau Voice

applicationselon-musk--x
23 Apr 2026
Tutorials

Transformers Can Learn Connectivity in Some Graphs but Not Others

DGX agent

arXiv:2509.22343v2 Announce Type: replace-cross Abstract: Reasoning capability is essential to ensure the factual correctness of the responses of transformer-based Large Language Models (LLMs), and ro

tutorialsarxiv-cs-ai
23 Apr 2026
Model Releases

a bunch here where I’m saying ok Garry’s kinda right?! 👀…in some ways :) we’re making this loop much easier to close out of the box soon If…

DGX agent

a bunch here where I’m saying ok Garry’s kinda right?! 👀…in some ways :) we’re making this loop much easier to close out of the box soon If more people get into evals & traces to ground self-improving

model-releasesharrison-chase--x
22 Apr 2026
Model Releases

AD-Copilot: A Vision-Language Assistant for Industrial Anomaly Detection via Visual In-context Comparison

DGX agent

arXiv:2603.13779v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have achieved impressive success in natural visual understanding, yet they consistently underperform

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System

DGX agent

arXiv:2604.18789v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) is central to aligning Large Language Models (LLMs), yet it introduces a critical vulnerability: an im

safetyarxiv-cs-ai
22 Apr 2026
Research

BED-LLM: Intelligent Information Gathering with LLMs and Bayesian Experimental Design

DGX agent

arXiv:2508.21184v3 Announce Type: replace-cross Abstract: We propose a general-purpose approach for improving the ability of large language models (LLMs) to intelligently and adaptively gather informa

researcharxiv-cs-ai
22 Apr 2026
Safety

Benchmarking Misuse Mitigation Against Covert Adversaries

DGX agent

arXiv:2506.06414v2 Announce Type: replace-cross Abstract: Existing language model safety evaluations focus on overt attacks and low-stakes tasks. In reality, an attacker can easily subvert existing sa

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Beyond Itinerary Planning-A Real-World Benchmark for Multi-Turn and Tool-Using Travel Tasks

DGX agent

arXiv:2512.22673v3 Announce Type: replace Abstract: Travel planning is a natural real-world task to test large language models' (LLMs) planning and tool-use abilities. Although prior work has studied

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Beyond Marginal Distributions: A Framework to Evaluate the Representativeness of Demographic-Aligned LLMs

DGX agent

arXiv:2601.15755v3 Announce Type: replace Abstract: Large language models are increasingly used to represent human opinions, values, or beliefs, and their steerability towards these ideals is an activ

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Characterizing AlphaEarth Embedding Geometry for Agentic Environmental Reasoning

DGX agent

arXiv:2604.18715v1 Announce Type: cross Abstract: Earth observation foundation models encode land surface information into dense embedding vectors, yet the geometric structure of these representations

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Day 1 at Google Cloud Next ‘26 recap

DGX agent

Last year at Google Cloud Next ‘25, we asked you to imagine a new future for AI. At Next ‘26, the question before you is how do you move AI into production across your entire enterprise? According to

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

Deep Supervised Contrastive Learning of Pitch Contours for Robust Pitch Accent Classification in Seoul Korean

DGX agent

arXiv:2604.19477v1 Announce Type: cross Abstract: The intonational structure of Seoul Korean has been defined with discrete tonal categories within the Autosegmental-Metrical model of intonational pho

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Do LLMs Game Formalization? Evaluating Faithfulness in Logical Reasoning

DGX agent

arXiv:2604.19459v1 Announce Type: new Abstract: Formal verification guarantees proof validity but not formalization faithfulness. For natural-language logical reasoning, where models construct axiom s

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Environmental Sound Deepfake Detection Using Deep-Learning Framework

DGX agent

arXiv:2604.19652v1 Announce Type: cross Abstract: In this paper, we propose a deep-learning framework for environmental sound deepfake detection (ESDD) -- the task of identifying whether the sound sce

model-releasesarxiv-cs-ai
22 Apr 2026
Local Ai

Evaluation-driven Scaling for Scientific Discovery

DGX agent

arXiv:2604.19341v1 Announce Type: cross Abstract: Language models are increasingly used in scientific discovery to generate hypotheses, propose candidate solutions, implement systems, and iteratively

local-aiarxiv-cs-ai
22 Apr 2026
Model Releases

Fine-tuning DeepSeek-OCR-2 for Molecular Structure Recognition

DGX agent

arXiv:2604.03476v2 Announce Type: replace-cross Abstract: Optical Chemical Structure Recognition (OCSR) is critical for converting 2D molecular diagrams from printed literature into machine-readable f

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

LLMs Know They're Wrong and Agree Anyway: The Shared Sycophancy-Lying Circuit

DGX agent

arXiv:2604.19117v1 Announce Type: new Abstract: When a language model agrees with a user's false belief, is it failing to detect the error, or noticing and agreeing anyway? We show the latter. Across

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

LSTM-MAS: A Long Short-Term Memory Inspired Multi-Agent System for Long-Context Understanding

DGX agent

arXiv:2601.11913v2 Announce Type: replace-cross Abstract: Effectively processing long contexts remains a fundamental yet unsolved challenge for large language models (LLMs). Existing single-LLM-based

model-releasesarxiv-cs-ai
22 Apr 2026
Tutorials

MapPFN: Learning Causal Perturbation Maps in Context

DGX agent

arXiv:2601.21092v2 Announce Type: replace Abstract: Planning effective interventions in biological systems requires treatment-effect models that adapt to unseen biological contexts by identifying thei

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Mechanistic Anomaly Detection via Functional Attribution

DGX agent

arXiv:2604.18970v1 Announce Type: new Abstract: We can often verify the correctness of neural network outputs using ground truth labels, but we cannot reliably determine whether the output was produce

model-releasesarxiv-cs-lg
22 Apr 2026
Model Releases

Multi-Domain Learning with Global Expert Mapping

DGX agent

arXiv:2604.18842v1 Announce Type: new Abstract: Human perception generalizes well across different domains, but most vision models struggle beyond their training data. This gap motivates multi-dataset

model-releasesarxiv-cs-cv
22 Apr 2026
Local Ai

Optimal Routing for Federated Learning over Dynamic Satellite Networks: Tractable or Not?

DGX agent

arXiv:2604.19399v1 Announce Type: new Abstract: Federated learning (FL) is a key paradigm for distributed model learning across decentralized data sources. Communication in each FL round typically con

local-aiarxiv-cs-lg
22 Apr 2026
Model Releases

PriorGuide: Test-Time Prior Adaptation for Simulation-Based Inference

DGX agent

arXiv:2510.13763v2 Announce Type: replace-cross Abstract: Amortized simulator-based inference offers a powerful framework for tackling Bayesian inference in computational fields such as engineering or

model-releasesarxiv-cs-lg
22 Apr 2026
Safety

Probing for Reading Times

DGX agent

arXiv:2604.18712v1 Announce Type: new Abstract: Probing has shown that language model representations encode rich linguistic information, but it remains unclear whether they also capture cognitive sig

safetyarxiv-cs-cl
22 Apr 2026
Model Releases

Safe Continual Reinforcement Learning in Non-stationary Environments

DGX agent

arXiv:2604.19737v1 Announce Type: new Abstract: Reinforcement learning (RL) offers a compelling data-driven paradigm for synthesizing controllers for complex systems when accurate physical models are

model-releasesarxiv-cs-lg
22 Apr 2026
Research

SimDiff: Depth Pruning via Similarity and Difference

DGX agent

arXiv:2604.19520v1 Announce Type: new Abstract: Depth pruning improves the deployment efficiency of large language models (LLMs) by identifying and removing redundant layers. A widely accepted standar

researcharxiv-cs-ai
22 Apr 2026
Model Releases

Taming Actor-Observer Asymmetry in Agents via Dialectical Alignment

DGX agent

arXiv:2604.19548v1 Announce Type: cross Abstract: Large Language Model agents have rapidly evolved from static text generators into dynamic systems capable of executing complex autonomous workflows. T

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Towards Reliable Human Evaluations in Gesture Generation: Insights from a Community-Driven State-of-the-Art Benchmark

DGX agent

arXiv:2511.01233v3 Announce Type: replace Abstract: We review human evaluation practices in automatic, speech-driven 3D gesture generation and find a lack of standardisation and frequent use of flawed

model-releasesarxiv-cs-cv
22 Apr 2026
Model Releases

Towards Scalable Lifelong Knowledge Editing with Selective Knowledge Suppression

DGX agent

arXiv:2604.19089v1 Announce Type: new Abstract: Large language models (LLMs) require frequent knowledge updates to reflect changing facts and mitigate hallucinations. To meet this demand, lifelong kno

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Unlocking the Edge deployment and ondevice acceleration of multi-LoRA enabled one-for-all foundational LLM

DGX agent

arXiv:2604.18655v1 Announce Type: cross Abstract: Deploying large language models (LLMs) on smartphones poses significant engineering challenges due to stringent constraints on memory, latency, and ru

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

What’s new in GKE at Next ‘26

DGX agent

This week at Google Cloud Next ‘26, we are sharing the evolution of Google Kubernetes Engine (GKE), delivering leading performance, efficiency, security, and scale for your most demanding and complex

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

Xpertbench: Expert Level Tasks with Rubrics-Based Evaluation

DGX agent

arXiv:2604.02368v4 Announce Type: replace Abstract: As Large Language Models (LLMs) exhibit plateauing performance on conventional benchmarks, a pivotal challenge persists: evaluating their proficienc

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

An Interpretable Framework Applying Protein Words to Predict Protein-Small Molecule Complementary Pairing Rules

DGX agent

arXiv:2604.16550v1 Announce Type: new Abstract: Despite the high accuracy of 'black box' deep learning models, drug discovery still relies on protein-ligand interaction principles and heuristics. To i

model-releasesarxiv-cs-lg
21 Apr 2026
Research

AutoRubric: Rubric-Based Generative Rewards for Faithful Multimodal Reasoning

DGX agent

arXiv:2510.14738v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have rapidly advanced from perception tasks to complex multi-step reasoning, yet reinforcement learning wit

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Balanced Co-Clustering of Users and Items for Embedding Table Compression in Recommender Systems

DGX agent

arXiv:2604.18351v1 Announce Type: cross Abstract: Recommender systems have advanced markedly over the past decade by transforming each user/item into a dense embedding vector with deep learning models

model-releasesarxiv-cs-lg
21 Apr 2026
Research

BioVLM: Routing Prompts, Not Parameters, for Cross-Modality Generalization in Biomedical VLMs

DGX agent

arXiv:2604.17629v1 Announce Type: new Abstract: Pretrained biomedical vision-language models (VLMs) such as BioMedCLIP perform well on average but often degrade on challenging modalities where inter-c

researcharxiv-cs-cv
21 Apr 2026
Model Releases

DriveAgent-R1: Advancing VLM-based Autonomous Driving with Active Perception and Hybrid Thinking

DGX agent

arXiv:2507.20879v3 Announce Type: replace Abstract: The advent of Vision-Language Models (VLMs) has significantly advanced end-to-end autonomous driving, demonstrating powerful reasoning abilities for

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

EditVerse: Unifying Image and Video Editing and Generation with In-Context Learning

DGX agent

arXiv:2509.20360v3 Announce Type: replace Abstract: Recent advances in foundation models highlight a clear trend toward unification and scaling, showing emergent capabilities across diverse domains. W

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

DGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

'Faithful to What?' On the Limits of Fidelity-Based Explanations

DGX agent

arXiv:2506.12176v5 Announce Type: replace Abstract: In explainable AI, surrogate models are commonly evaluated by their fidelity to a neural network's predictions. Fidelity, however, measures alignmen

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

FedOBP: Federated Optimal Brain Personalization through Cloud-Edge Element-wise Decoupling

DGX agent

arXiv:2604.16574v1 Announce Type: new Abstract: Federated Learning (FL) faces challenges from client data heterogeneity and resource-constrained mobile devices, which can degrade model accuracy. Perso

model-releasesarxiv-cs-lg
21 Apr 2026
Tutorials

From Adaptation to Generalization: Adaptive Visual Prompting for Medical Image Segmentation

DGX agent

arXiv:2604.17455v1 Announce Type: new Abstract: Visual prompting has emerged as a powerful method for adapting pre-trained models to new domains without updating model parameters. However, existing pr

tutorialsarxiv-cs-cv
21 Apr 2026
Model Releases

From Inheritance to Saturation: Disentangling the Evolution of Visual Redundancy for Architecture-Aware MLLM Inference Acceleration

DGX agent

arXiv:2604.16462v1 Announce Type: new Abstract: High-resolution Multimodal Large Language Models (MLLMs) face prohibitive computational costs during inference due to the explosion of visual tokens. Ex

model-releasesarxiv-cs-cv
21 Apr 2026
← Previous
1…408409410411412…1326
Next →