AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,288 results
Model Releases

MemOCR: Layout-Aware Visual Memory for Efficient Long-Horizon Reasoning

DGX agent

arXiv:2601.21468v5 Announce Type: replace Abstract: Long-horizon agentic reasoning necessitates effectively compressing growing interaction histories into a limited context window. Most existing memor

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

MentalBench: A DSM-Grounded Benchmark for Evaluating Psychiatric Diagnostic Capability of Large Language Models

DGX agent

arXiv:2602.12871v2 Announce Type: replace Abstract: Large language models (LLMs) have attracted growing interest as supportive tools for psychiatric assessment and clinical decision support. However,

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Merlin's Whisper: Enabling Efficient Reasoning in Large Language Models via Black-box Persuasive Prompting

DGX agent

arXiv:2510.10528v3 Announce Type: replace Abstract: Large reasoning models (LRMs) have demonstrated remarkable proficiency in tackling complex tasks through step-by-step thinking. However, this length

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

DGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MiniGPT: Rebuilding GPT from First Principles

DGX agent

arXiv:2605.17398v1 Announce Type: new Abstract: This paper presents MiniGPT, a compact from-scratch implementation of GPT-style autoregressive language modeling in PyTorch. The aim is to rebuild the c

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MIRAGE: Robust multi-modal architectures translate fMRI-to-image models from vision to mental imagery

DGX agent

arXiv:2605.17198v1 Announce Type: cross Abstract: To be useful for downstream applications, vision decoding models that are trained to reconstruct seen images from human brain activity must be able to

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

MirrorBench: A Benchmark to Evaluate Conversational User-Proxy Agents for Human-Likeness

DGX agent

arXiv:2601.08118v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as human simulators, both for evaluating conversational systems and for generating fine-tuning da

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MixSD: Mixed Contextual Self-Distillation for Knowledge Injection

DGX agent

arXiv:2605.16865v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) is widely used to inject new knowledge into language models, but it often degrades pretrained capabilities such as reasonin

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Mixture-of-Experts Can Surpass Dense LLMs Under Strictly Equal Resource

DGX agent

arXiv:2506.12119v2 Announce Type: replace-cross Abstract: Mixture-of-Experts (MoE) language models dramatically expand model capacity and achieve remarkable performance without increasing per-token co

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Mixture of Experts for Low-Resource LLMs

DGX agent

arXiv:2605.17598v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) architectures enable efficient model scaling, yet expert routing behavior across underrepresented languages remains poorly unde

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

MLReplicate: Benchmarking Autonomous Research Systems for Machine Learning Reproducibility

DGX agent

arXiv:2605.16616v1 Announce Type: new Abstract: Autonomous research systems capable of generating complete scientific manuscripts have advanced rapidly, yet robust and realistic evaluation frameworks

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Monocular Open Vocabulary Occupancy Prediction for Indoor Scenes

DGX agent

arXiv:2602.22667v2 Announce Type: replace Abstract: Open-vocabulary 3D occupancy is vital for embodied agents, which need to understand complex indoor environments where semantic categories are abunda

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

MorphSeek: Fine-grained Latent Representation-Level Policy Optimization for Deformable Image Registration

DGX agent

arXiv:2511.17392v3 Announce Type: replace Abstract: Deformable image registration (DIR) remains a fundamental yet challenging problem in medical image analysis, largely due to the prohibitively high-d

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination

DGX agent

arXiv:2605.17454v1 Announce Type: new Abstract: Multi-party multi-objective optimization problems (MPMOPs) require consensus among autonomous decision makers and therefore differ from flattened many-o

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Multi-site PPG: An In-the-Wild Physiological Dataset from Emerging Multi-site Wearables

DGX agent

arXiv:2605.17859v1 Announce Type: cross Abstract: Wearables are widely used for mobile health monitoring, and photoplethysmography (PPG) is a key sensing modality for heart rate and related physiologi

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Multilingual jailbreaking of LLMs using low-resource languages

DGX agent

arXiv:2605.18239v1 Announce Type: cross Abstract: Large Language Models (LLMs) remain vulnerable to jailbreak attempts that circumvent safety guardrails. We investigate whether multi-turn conversation

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models

DGX agent

arXiv:2605.16409v1 Announce Type: cross Abstract: Optical character recognition (OCR) and multilingual text understanding remain major failure modes of multimodal large language models (MLLMs), partic

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

DGX agent

arXiv:2605.17669v1 Announce Type: new Abstract: The preservation and interpretation of cultural heritage increasingly rely on digital technologies, among which Knowledge Graphs (KGs) stand out for the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

NeuroMAS: Multi-Agent Systems as Neural Networks with Joint Reinforcement Learning

DGX agent

arXiv:2605.16757v1 Announce Type: new Abstract: Multi-agent language systems are often built as hand-designed workflows, where agents are assigned semantic roles and communication protocols are specif

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Neuroscience-inspired Staged Representation Learning with Disentangled Coarse- and Fine-Grained Semantics for EEG Visual Decoding

DGX agent

arXiv:2605.16923v1 Announce Type: new Abstract: Decoding visual information from electroencephalography (EEG) signals remains a fundamental challenge in brain-computer interfaces and medical rehabilit

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

DGX agent

arXiv:2605.17596v1 Announce Type: new Abstract: We present NeuSymMS, an adaptive memory system that enables large language model (LLM) agents to learn, remember, and reason about users across sessions

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

NewsLens: A Multi-Agent Framework for Adversarial News Bias Navigation

DGX agent

arXiv:2605.17364v1 Announce Type: new Abstract: Media bias detection has predominantly been framed as a classification task: assign a political label to an article or outlet. We argue this framing is

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Noise2Params: Unification and Parameter Determination from Noise via a Probabilistic Event Camera Model

DGX agent

arXiv:2605.16317v1 Announce Type: new Abstract: Accurate, unified models for event cameras (ECs) remain elusive, hampering calibration and algorithm design. We develop a foundational probabilistic mod

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

Nori Bot: A Sub-$1,000 Floor-to-Counter Mobile Manipulator

DGX agent

arXiv:2605.16537v1 Announce Type: new Abstract: Open-source mobile manipulators have reached 660 (XLeRobot) but every sub-1,000 platform shares three limitations: a fixed-height workspace, reactive-on

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

Not What You Asked For: Typographic Attacks in Household Robot Manipulation

DGX agent

arXiv:2605.18593v1 Announce Type: cross Abstract: Open-vocabulary embodied AI agents increasingly rely on vision-language models such as CLIP for object perception and task grounding. However, the sha

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Offline Contextual Bandits in the Presence of New Actions

DGX agent

arXiv:2605.18509v1 Announce Type: new Abstract: Automated decision-making algorithms drive applications such as recommendation systems and search engines. These algorithms often rely on off-policy con

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Omni-DuplexEval: Evaluating Real-time Duplex Omni-modal Interaction

DGX agent

arXiv:2605.17360v1 Announce Type: new Abstract: Real-time duplex interaction is essential for multimodal AI systems operating in real-world scenarios, where models must continuously process streaming

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

OmniCode: A Benchmark for Evaluating Software Engineering Agents

DGX agent

arXiv:2602.02262v3 Announce Type: replace-cross Abstract: LLM-powered coding agents are redefining how real-world software is developed. To drive the research towards better coding agents, we require

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

OmniPro: A Comprehensive Benchmark for Omni-Proactive Streaming Video Understanding

DGX agent

arXiv:2605.18577v1 Announce Type: new Abstract: Omni-proactive streaming video understanding, i.e., autonomously deciding when to speak and what to say from continuous audio-visual streams, is an emer

model-releasesarxiv-cs-cv
19 May 2026
Model Releases

OmniVL-Guard Pro: A Tool-Augmented Agent for Omnibus Vision-Language Forensics

DGX agent

arXiv:2605.16962v1 Announce Type: cross Abstract: Existing vision-language forgery detection and grounding methods operate under a closed-world paradigm, assuming verification can be completed by the

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

On Improving Multimodal Pedestrian Trajectory Prediction with CVAE: A Study on Benchmark and Robot Data

DGX agent

arXiv:2605.18262v1 Announce Type: new Abstract: Accurate pedestrian trajectory prediction is crucial for autonomous systems operating in complex environments, such as modular buses and delivery robots

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

One Hand to Rule Them All: Canonical Representations for Unified Dexterous Manipulation

DGX agent

arXiv:2602.16712v2 Announce Type: replace Abstract: Dexterous manipulation policies today largely assume fixed hand designs, severely restricting their generalization to new embodiments with varied ki

model-releasesarxiv-cs-ro
19 May 2026
Model Releases

One Model, Two Roles: Emergent Specialization in a Shared Recurrent Transformer

DGX agent

arXiv:2605.17811v1 Announce Type: cross Abstract: Can a shared-weight recurrent Transformer develop distinct internal roles without being partitioned into separate modules? We study this in Asymmetric

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Online Resource Allocation with Convex-set Machine-Learned Advice

DGX agent

arXiv:2306.12282v2 Announce Type: replace-cross Abstract: Decision-makers often have access to machine-learned predictions about future demand that can help guide online resource allocation decisions.

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

OpenJarvis: Personal AI, On Personal Devices

DGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Optimising CSRNet with parameter-free attention mechanisms for crowd counting in public transport

DGX agent

arXiv:2605.18349v1 Announce Type: cross Abstract: Occupancy estimation and crowd counting are critical tasks in designing smart and efficient public transport vehicles. Given that public transport loa

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ORACLE: Anticipating Scams from Partial Trajectories in Streaming App Usage

DGX agent

arXiv:2605.16363v1 Announce Type: new Abstract: Smartphone scams are increasingly prevalent and typically manifest as multi-stage, cross-application processes with gradually emerging intent. Effective

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

Ordinal Adaptive Correction: A Data-Centric Approach to Ordinal Image Classification with Noisy Labels

DGX agent

arXiv:2509.02351v3 Announce Type: replace-cross Abstract: Labeled data is a fundamental component in training supervised deep learning models for computer vision tasks. However, the labeling process,

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

OSWorld-Human: Benchmarking the Efficiency of Computer-Use Agents

DGX agent

arXiv:2506.16042v2 Announce Type: replace Abstract: Generative AI is being leveraged to solve a variety of computer-use tasks involving desktop applications. State-of-the-art systems have focused sole

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

DGX agent

arXiv:2605.18583v1 Announce Type: cross Abstract: Coding agents now run autonomously with shell, file, and network privileges. When a user issues a benign request, the agent sometimes does more than a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PACE: Geometry-Aware Bridge Transport for Single-Cell Trajectory Inference

DGX agent

arXiv:2605.18587v1 Announce Type: cross Abstract: Single-cell trajectory inference from destructive time-course snapshots is fundamentally ill-posed: neither cross-time cell correspondences nor contin

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

PaliBench: A Multi-Reference Blueprint for Classical Language Translation Benchmarks

DGX agent

arXiv:2605.16881v1 Announce Type: new Abstract: Digital humanities projects increasingly rely on machine translation and large language models to widen access to classical, religious, and otherwise un

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

PARALLAX: Separating Genuine Hallucination Detection from Benchmark Construction Artifacts

DGX agent

arXiv:2605.17028v1 Announce Type: cross Abstract: Large language models (LLMs) hallucinate with confidence: their outputs can be fluent, authoritative, and simply wrong. In medical, legal, and scienti

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Parameter-Efficient Domain Adaptation of Physics-Informed Self-Attention based GNNs for AC Power Flow Prediction

DGX agent

arXiv:2602.18227v2 Announce Type: replace Abstract: Accurate AC power flow (AC-PF) prediction under domain shift is critical when models trained on medium-voltage (MV) grids are deployed on high-volta

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

PAREDA: A Multi-Accent Speech Dataset of Natural Language Processing Research Discussions

DGX agent

arXiv:2605.17860v1 Announce Type: cross Abstract: While modern Automatic Speech Recognition (ASR) systems achieve high accuracy on benchmark corpora, their performance often degrades when there is rea

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

PEGRL: Improving Machine Translation by Post-Editing Guided Reinforcement Learning

DGX agent

arXiv:2602.03352v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown strong promise for LLM-based machine translation, with recent methods such as GRPO demonstrating notable gains

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

Perfect Parallelization in Mini-Batch SGD with Classical Momentum Acceleration

DGX agent

arXiv:2605.18609v1 Announce Type: new Abstract: Accelerating stochastic gradient methods with classical momentum schemes, such as Polyak's heavy ball, has proven highly successful in training large-sc

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

PERL: Parameter Efficient Reasoning in CLIP Latent Space

DGX agent

arXiv:2605.18464v1 Announce Type: new Abstract: Contrastively trained vision-language models such as CLIP provide strong zero-shot transfer by aligning images and text in a shared embedding space. How

model-releasesarxiv-cs-cv
19 May 2026
← Previous
1…227228229230231…361
Next →