AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Breaking the Generator Barrier: Disentangled Representation for Generalizable AI-Text Detection

DGX agent

arXiv:2604.13692v1 Announce Type: new Abstract: As large language models (LLMs) generate text that increasingly resembles human writing, the subtle cues that distinguish AI-generated content from huma

model-releasesarxiv-cs-cl
16 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

DGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Can Large Language Models Reliably Extract Physiology Index Values from Coronary Angiography Reports?

DGX agent

arXiv:2604.13077v1 Announce Type: new Abstract: Coronary angiography (CAG) reports contain clinically relevant physiological measurements, yet this information is typically in the form of unstructured

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CANVAS: Continuity-Aware Narratives via Visual Agentic Storyboarding

DGX agent

arXiv:2604.13452v1 Announce Type: new Abstract: Long-form visual storytelling requires maintaining continuity across shots, including consistent characters, stable environments, and smooth scene trans

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Chain of Uncertain Rewards with Large Language Models for Reinforcement Learning

DGX agent

arXiv:2604.13504v1 Announce Type: cross Abstract: Designing effective reward functions is a cornerstone of reinforcement learning (RL), yet it remains a challenging and labor-intensive process due to

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CodeFlowBench: A Multi-turn, Iterative Benchmark for Complex Code Generation

DGX agent

arXiv:2504.21751v4 Announce Type: replace-cross Abstract: Modern software development demands code that is maintainable, testable, and scalable by organizing the implementation into modular components

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Coherence in the brain unfolds across separable temporal regimes

DGX agent

arXiv:2512.20481v4 Announce Type: replace-cross Abstract: To maintain coherence in language, the brain must satisfy key competing temporal demands: the gradual accumulation of meaning across extended

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation

DGX agent

arXiv:2604.13946v1 Announce Type: cross Abstract: Automated code generation remains a persistent challenge in software engineering, as conventional multi-agent frameworks are often constrained by stat

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Common to Whom? Regional Cultural Commonsense and LLM Bias in India

DGX agent

arXiv:2601.15550v3 Announce Type: replace Abstract: Existing cultural commonsense benchmarks treat nations as monolithic, assuming uniform practices within national boundaries. But does cultural commo

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Correct Chains, Wrong Answers: Dissociating Reasoning from Output in LLM Logic

DGX agent

arXiv:2604.13065v1 Announce Type: new Abstract: LLMs can execute every step of chain-of-thought reasoning correctly and still produce wrong final answers. We introduce the Novel Operator Test, a bench

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Correct Prediction, Wrong Steps? Consensus Reasoning Knowledge Graph for Robust Chain-of-Thought Synthesis

DGX agent

arXiv:2604.14121v1 Announce Type: new Abstract: LLM reasoning traces suffer from complex flaws -- *Step Internal Flaws* (logical errors, hallucinations, etc.) and *Step-wise Flaws* (overthinking, unde

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Counterfactual Peptide Editing for Causal TCR--pMHC Binding Inference

DGX agent

arXiv:2604.13256v1 Announce Type: new Abstract: Neural models for TCR-pMHC binding prediction are susceptible to shortcut learning: they exploit spurious correlations in training data -- such as pepti

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Covariance-adapting algorithm for semi-bandits with application to sparse rewards

DGX agent

arXiv:2604.13738v1 Announce Type: cross Abstract: We investigate stochastic combinatorial semi-bandits, where the entire joint distribution of outcomes impacts the complexity of the problem instance (

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Cracking the Code of Juxtaposition: Can AI Models Understand the Humorous Contradictions

DGX agent

arXiv:2405.19088v3 Announce Type: replace Abstract: Recent advancements in large multimodal language models have demonstrated remarkable proficiency across a wide range of tasks. Yet, these models sti

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Data-driven Learning of Probabilistic Model of Binary Droplet Collision for Spray Simulation

DGX agent

arXiv:2604.13594v1 Announce Type: cross Abstract: Binary droplet collisions are ubiquitous in dense sprays. Traditional deterministic models cannot adequately represent transitional and stochastic beh

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Decoding the Delta: Unifying Remote Sensing Change Detection and Understanding with Multimodal Large Language Models

DGX agent

arXiv:2604.14044v1 Announce Type: new Abstract: While Multimodal Large Language Models (MLLMs) excel in general vision-language tasks, their application to remote sensing change understanding is hinde

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

DeEscalWild: A Real-World Benchmark for Automated De-Escalation Training with SLMs

DGX agent

arXiv:2604.13075v1 Announce Type: new Abstract: Effective de-escalation is critical for law enforcement safety and community trust, yet traditional training methods lack scalability and realism. While

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage

DGX agent

arXiv:2604.13060v1 Announce Type: new Abstract: Dental triage is a safety-critical clinical routing task that requires integrating multimodal clinical information (e.g., patient complaints and radiogr

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Design Space Exploration of Hybrid Quantum Neural Networks for Chronic Kidney Disease

DGX agent

arXiv:2604.13608v1 Announce Type: new Abstract: Hybrid Quantum Neural Networks (HQNNs) have recently emerged as a promising paradigm for near-term quantum machine learning. However, their practical pe

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

DF3DV-1K: A Large-Scale Dataset and Benchmark for Distractor-Free Novel View Synthesis

DGX agent

arXiv:2604.13416v1 Announce Type: new Abstract: Advances in radiance fields have enabled photorealistic novel view synthesis. In several domains, large-scale real-world datasets have been developed to

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Do We Still Need Humans in the Loop? Comparing Human and LLM Annotation in Active Learning for Hostility Detection

DGX agent

arXiv:2604.13899v1 Announce Type: new Abstract: Instruction-tuned LLMs can annotate thousands of instances from a short prompt at negligible cost. This raises two questions for active learning (AL): c

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Document-tuning for robust alignment to animals

DGX agent

arXiv:2604.13076v1 Announce Type: new Abstract: We investigate the robustness of value alignment via finetuning with synthetic documents, using animal compassion as a value that is both important in i

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Drowsiness-Aware Adaptive Autonomous Braking System based on Deep Reinforcement Learning for Enhanced Road Safety

DGX agent

arXiv:2604.13878v1 Announce Type: new Abstract: Driver drowsiness significantly impairs the ability to accurately judge safe braking distances and is estimated to contribute to 10%-20% of road acciden

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Efficient Multi-View 3D Object Detection by Dynamic Token Selection and Fine-Tuning

DGX agent

arXiv:2604.13586v1 Announce Type: new Abstract: Existing multi-view three-dimensional (3D) object detection approaches widely adopt large-scale pre-trained vision transformer (ViT)-based foundation mo

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

EmbodiedClaw: Conversational Workflow Execution for Embodied AI Development

DGX agent

arXiv:2604.13800v1 Announce Type: new Abstract: Embodied AI research is increasingly moving beyond single-task, single-environment policy learning toward multi-task, multi-scene, and multi-model setti

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

EMGFlow: Robust and Efficient Surface Electromyography Synthesis via Flow Matching

DGX agent

arXiv:2604.13685v1 Announce Type: cross Abstract: Deep learning-based surface electromyography (sEMG) gesture recognition is frequently bottlenecked by data scarcity and limited subject diversity. Whi

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Enhancing Confidence Estimation in Telco LLMs via Twin-Pass CoT-Ensembling

DGX agent

arXiv:2604.13271v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly applied to complex telecommunications tasks, including 3GPP specification analysis and O-RAN network troub

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

ESCAPE: Episodic Spatial Memory and Adaptive Execution Policy for Long-Horizon Mobile Manipulation

DGX agent

arXiv:2604.13633v1 Announce Type: new Abstract: Coordinating navigation and manipulation with robust performance is essential for embodied AI in complex indoor environments. However, as tasks extend o

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Evaluating LLM-Based Translation of a Low-Resource Technical Language: The Medical and Philosophical Greek of Galen

DGX agent

arXiv:2602.24119v2 Announce Type: replace Abstract: Purpose: This study evaluates the quality of commercial large language model (LLM) machine translation (MT) for Ancient Greek technical prose and be

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Evaluating Supervised Machine Learning Models: Principles, Pitfalls, and Metric Selection

DGX agent

arXiv:2604.13882v1 Announce Type: new Abstract: The evaluation of supervised machine learning models is a critical stage in the development of reliable predictive systems. Despite the widespread avail

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Evaluating the Evaluator: Problems with SemEval-2020 Task 1 for Lexical Semantic Change Detection

DGX agent

arXiv:2604.13232v1 Announce Type: new Abstract: This discussion paper re-examines SemEval-2020 Task 1, the most influential shared benchmark for lexical semantic change detection, through a three-part

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy

DGX agent

arXiv:2604.02709v2 Announce Type: replace Abstract: The formal reasoning capabilities of LLMs are crucial for advancing automated software engineering. However, existing benchmarks for LLMs lack syste

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

EVE: A Domain-Specific LLM Framework for Earth Intelligence

DGX agent

arXiv:2604.13071v1 Announce Type: new Abstract: We introduce Earth Virtual Expert (EVE), the first open-source, end-to-end initiative for developing and deploying domain-specialized LLMs for Earth Int

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Exposia: Teaching and Assessment of Academic Writing Skills for Research Project Proposals and Peer Feedback

DGX agent

arXiv:2601.06536v2 Announce Type: replace Abstract: We present Exposia, the first public dataset that connects writing and feedback in higher education, enabling research on educationally grounded com

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

ExpSeek: Self-Triggered Experience Seeking for Web Agents

DGX agent

arXiv:2601.08605v2 Announce Type: replace Abstract: Experience intervention in web agents emerges as a promising technical paradigm, enhancing agent interaction capabilities by providing valuable insi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

F-Actor: Controllable Conversational Behaviour in Full-Duplex Models

DGX agent

arXiv:2601.11329v3 Announce Type: replace Abstract: Spoken conversational systems require more than accurate speech generation to have human-like conversations: to feel natural and engaging, they must

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Failure Makes the Agent Stronger: Enhancing Accuracy through Structured Reflection for Reliable Tool Interactions

DGX agent

arXiv:2509.18847v3 Announce Type: replace-cross Abstract: Tool-augmented large language models (LLMs) are usually trained with supervised imitation or coarse-grained reinforcement learning that optimi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Fast training of accurate physics-informed neural networks without gradient descent

DGX agent

arXiv:2405.20836v3 Announce Type: replace-cross Abstract: Solving time-dependent Partial Differential Equations (PDEs) is one of the most critical problems in computational science. While Physics-Info

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

FieldWorkArena: Agentic AI Benchmark for Real Field Work Tasks

DGX agent

arXiv:2505.19662v3 Announce Type: replace-cross Abstract: This paper introduces FieldWorkArena, a benchmark for agentic AI targeting real-world field work. With the recent increase in demand for agent

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

FiLM-Nav: Efficient and Generalizable Navigation via VLM Fine-tuning

DGX agent

arXiv:2509.16445v2 Announce Type: replace Abstract: Enabling robotic assistants to navigate complex environments and locate objects described in free-form language is a critical capability for real-wo

model-releasesarxiv-cs-ro
16 Apr 2026
Model Releases

FlexGuard: Continuous Risk Scoring for Strictness-Adaptive LLM Content Moderation

DGX agent

arXiv:2602.23636v3 Announce Type: replace Abstract: Ensuring the safety of LLM-generated content is essential for real-world deployment. Most existing guardrail models formulate moderation as a fixed

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

Flow-based Generative Modeling of Potential Outcomes and Counterfactuals

DGX agent

arXiv:2505.16051v4 Announce Type: replace-cross Abstract: Predicting potential and counterfactual outcomes from observational data is central to individualized decision-making, particularly in clinica

model-releasesarxiv-cs-lg
16 Apr 2026
Model Releases

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

DGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Free Geometry: Refining 3D Reconstruction from Longer Versions of Itself

DGX agent

arXiv:2604.14048v1 Announce Type: new Abstract: Feed-forward 3D reconstruction models are efficient but rigid: once trained, they perform inference in a zero-shot manner and cannot adapt to the test s

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

From Feelings to Metrics: Understanding and Formalizing How Users Vibe-Test LLMs

DGX agent

arXiv:2604.14137v1 Announce Type: new Abstract: Evaluating LLMs is challenging, as benchmark scores often fail to capture models' real-world usefulness. Instead, users often rely on ``vibe-testing'':

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

From Plausibility to Verifiability: Risk-Controlled Generative OCR with Vision-Language Models

DGX agent

arXiv:2603.19790v3 Announce Type: replace Abstract: Modern vision-language models (VLMs) can act as generative OCR engines, yet open-ended decoding can expose rare but consequential failures. We ident

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

From Weights to Activations: Is Steering the Next Frontier of Adaptation?

DGX agent

arXiv:2604.14090v1 Announce Type: new Abstract: Post-training adaptation of language models is commonly achieved through parameter updates or input-based methods such as fine-tuning, parameter-efficie

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Functional Emotions or Situational Contexts? A Discriminating Test from the Mythos Preview System Card

DGX agent

arXiv:2604.13466v1 Announce Type: cross Abstract: The Claude Mythos Preview system card deploys emotion vectors, sparse autoencoder (SAE) features, and activation verbalisers to study model internals

model-releasesarxiv-cs-cl
16 Apr 2026
← Previous
1…328329330331332…357
Next →