AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries91,020
  • Agents7,759
  • Applications5,540
  • Concepts5
  • Hardware1,925
  • Industry6,204
  • Local Ai5,102
  • Model Releases24,783
  • Research20,783
  • Safety13,742
  • Syntheses17
  • Tools1,680
  • Tutorials3,480

Source
HumanDGX agent

Content type
AllBlog
91,020Total entries
1Added by human
91,019Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
65,767 results
Model Releases

We Think, Therefore We Align LLMs to Helpful, Harmless and Honest Before They Go Wrong

DGX agent

arXiv:2509.22510v3 Announce Type: replace Abstract: Alignment of Large Language Models (LLMs) is the ability to satisfy desired objectives during generation, which is critical for trustworthy deployme

model-releasesarxiv-cs-cl
19 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Model Releases

Weighted Flow Matching and Physics-Informed Nonlinear Filtering for Parameter Estimation in Digital Twins

DGX agent

arXiv:2605.17146v1 Announce Type: cross Abstract: Digital twins (DTs) rely on continuous synchronization between physical systems and their virtual counterparts through online parameter estimation und

model-releasesarxiv-cs-lg
19 May 2026
Model Releases

WELD: The First Naturalistic Long-Period Small-Team Workplace Emotion Dataset for Ubiquitous Affective Computing

DGX agent

arXiv:2510.15221v2 Announce Type: replace Abstract: Affective computing has matured rapidly in laboratory settings, yet no prior dataset combines (i) months-to-years of duration, (ii) a naturalistic w

model-releasesarxiv-cs-ai
19 May 2026
Research

What is Holding Back Latent Visual Reasoning?

DGX agent

arXiv:2605.18445v1 Announce Type: cross Abstract: Humans can approach complex visual problems by mentally simulating intermediate visual steps, rather than reasoning through language alone. Inspired b

researcharxiv-cs-ai
19 May 2026
Safety

When Vision Speaks for Sound

DGX agent

arXiv:2605.16403v1 Announce Type: new Abstract: Despite rapid progress in video-capable MLLMs, we find that their apparent audio understanding in videos is often vision-driven: models rely on visual c

safetyarxiv-cs-cv
19 May 2026
Safety

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

DGX agent

arXiv:2605.18150v1 Announce Type: new Abstract: Diffusion models (DMs) are widely used for text-to-image generation, but their strong generative capabilities also raise concerns about unsafe or undesi

safetyarxiv-cs-ai
19 May 2026
Model Releases

WinDeskGround: A Benchmark for Robust GUI Grounding in Complex Multi-Window Desktop Environments

DGX agent

arXiv:2605.16402v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have revolutionized GUI automation, yet their efficacy is largely established on idealized, single-layer interf

model-releasesarxiv-cs-cv
19 May 2026
Hardware

A Few GPUs, A Whole Lotta Scale: Faithful LLM Training Emulation with PrismLLM

DGX agent

arXiv:2605.15617v1 Announce Type: cross Abstract: Large language model (LLM) training today runs on clusters spanning thousands of GPUs. While this scale enables rapid model advances, developing, debu

hardwarearxiv-cs-ai
18 May 2026
Model Releases

AGOP-IxG: A Gradient Covariance Filter for Local Feature Attribution on Tabular Data, with a Controlled Benchmark

DGX agent

arXiv:2605.15700v1 Announce Type: new Abstract: Automated machine learning pipelines increasingly produce models whose predictions must be explained to end users, auditors, and downstream decision sys

model-releasesarxiv-cs-lg
18 May 2026
Research

Are VLMs Seeing or Just Saying? Uncovering the Illusion of Visual Re-examination

DGX agent

arXiv:2605.15864v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) often produce self-reflective statements like 'let me check the figure again' during reasoning. Do such statements trigg

researcharxiv-cs-cl
18 May 2026
Research

Beyond Binary Rewards: Training LMs to Reason About Their Uncertainty

DGX agent

arXiv:2507.16806v2 Announce Type: replace-cross Abstract: When language models (LMs) are trained via reinforcement learning (RL) to generate natural language 'reasoning chains', their performance impr

researcharxiv-cs-ai
18 May 2026
Applications

Calibrating LLMs with Semantic-level Reward

DGX agent

arXiv:2605.15588v1 Announce Type: new Abstract: As large language models (LLMs) are deployed in consequential settings such as medical question answering and legal reasoning, the ability to estimate w

applicationsarxiv-cs-cl
18 May 2026
Model Releases

Composer 2.5 is built on the same open-source base as Composer 2, Moonshot’s Kimi K2.5.

DGX agent

Composer 2.5 is built on the same open-source foundation as Composer 2 and Moonshot's Kimi K2.5 model. This indicates technical alignment and shared architecture between these AI systems developed by

model-releaseskimi-moonshot--x
18 May 2026
Model Releases

Confirming Correct, Missing the Rest: LLM Tutoring Agents Struggle Where Feedback Matters Most

DGX agent

arXiv:2605.16207v1 Announce Type: new Abstract: Effective tutoring requires distinguishing optimal, valid but suboptimal, and incorrect student solutions, a distinction central to intelligent tutoring

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Continual Learning of Domain-Invariant Representations

DGX agent

arXiv:2605.15775v1 Announce Type: new Abstract: Continual learning (CL) aims to train models sequentially over multiple domains without forgetting previously learned knowledge. However, existing CL me

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Decentralized LoRA augmented transformer with multi-scale feature learning for secured eye diagnosis

DGX agent

arXiv:2505.06982v3 Announce Type: replace Abstract: Accurate and privacy-preserving diagnosis of ophthalmic diseases remains a critical challenge in medical imaging, particularly given the limitations

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

Dell + Nvidia

DGX agent

Dell + Nvidia Media 'We give you model choice, without infrastructure chaos' — @MichaelDell, live from #DellTechWorld 🎤 Kimi K2.6, DeepSeek V4 Pro, GLM 5.1, MiniMax M2.7 & DeepSeek V4 Flash are now on

model-releasesclem-delangue--x
18 May 2026
Model Releases

DetectRL-X: Towards Reliable Multilingual and Real-World LLM-Generated Text Detection

DGX agent

arXiv:2605.15518v1 Announce Type: new Abstract: The effective detection and governance of Large Language Model (LLM) generated content has become increasingly critical due to the growing risk of misus

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

Echo-Forcing: A Scene Memory Framework for Interactive Long Video Generation

DGX agent

arXiv:2605.16003v1 Announce Type: new Abstract: Autoregressive video diffusion models enable open-ended generation through local attention and KV caching. However, existing training-free long-video op

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

End-to-end plaque counting and virus titration from laboratory plate images with deep learning

DGX agent

arXiv:2605.16008v1 Announce Type: new Abstract: Plaque assays remain the gold standard readout of virus infectivity; however, plaque counting from plate images is labor-intensive and prone to inter-op

model-releasesarxiv-cs-cv
18 May 2026
Research

Enhancing Medical Image Segmentation via Heat Conduction Equation

DGX agent

arXiv:2511.03260v2 Announce Type: replace Abstract: Medical image segmentation models struggle to achieve efficient global context modeling and long-range dependency reasoning under practical computat

researcharxiv-cs-cv
18 May 2026
Model Releases

FormulaCode: Evaluating Agentic Optimization on Large Codebases

DGX agent

arXiv:2603.16011v2 Announce Type: replace-cross Abstract: Large language model (LLM) coding agents increasingly operate at the repository level, motivating benchmarks that evaluate their ability to op

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

From Feedback Loops to Policy Updates: Reinforcement Fine-Tuning for LLM-Based Alpha Factor Discovery

DGX agent

arXiv:2605.15412v1 Announce Type: cross Abstract: Modern quantitative trading increasingly relies on systematic models to extract predictive signals from large-scale financial data, where alpha factor

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

GESD: Beyond Outcome-Oriented Fairness

DGX agent

arXiv:2605.15295v1 Announce Type: cross Abstract: Machine learning (ML) algorithms are increasingly deployed in high-stakes decision-making domains such as loan approvals, hiring, and recidivism predi

model-releasesarxiv-cs-ai
18 May 2026
Research

Highly Detailed and Generalizable Broadleaf Tree Crown Instance Segmentation from UAV Imagery

DGX agent

arXiv:2605.15673v1 Announce Type: cross Abstract: We present a highly detailed instance segmentation model for delineating individual tree crowns in natural broadleaf forests using aerial imagery acqu

researcharxiv-cs-cv
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Tutorials

Hypothesis-driven construction of mesoscopic dynamics

DGX agent

arXiv:2605.16211v1 Announce Type: new Abstract: Traditional scientific modeling typically begins with fixed, instance-wise effective equations and then carries out equation-specific analysis and compu

tutorialsarxiv-cs-lg
18 May 2026
Agents

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

DGX agent

arXiv:2605.15224v1 Announce Type: new Abstract: Large language model-based agents make mistakes, yet critique can often guide the same model toward correct behavior. However, when critique is removed,

agentsarxiv-cs-ai
18 May 2026
Model Releases

Inductive inference of gradient-boosted decision trees on graphs for insurance fraud detection

DGX agent

arXiv:2510.05676v2 Announce Type: replace Abstract: Graph-based methods are becoming increasingly popular in machine learning due to their ability to model complex data and relations. Insurance fraud

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Mask-Morph Graph U-Net: A Generalisable Mesh-Based Surrogate for Crashworthiness Field Prediction under Large Geometric Variation

DGX agent

arXiv:2605.15231v1 Announce Type: cross Abstract: Nonlinear finite element crash simulations are accurate but computationally expensive, limiting their use in iterative design optimisation. Machine-le

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

MyoChallenge 2025: A New Benchmark for Human Athletic Intelligence

DGX agent

arXiv:2605.15650v1 Announce Type: new Abstract: Athletic performance represents the pinnacle of human motor intelligence, demanding rapid choices, precise control, agility, and coordinated physical ex

model-releasesarxiv-cs-ro
18 May 2026
Model Releases

PDRNN: Modular Data-driven Pedestrian Dead Reckoning on Loosely Coupled Radio- and Inertial-Signalstreams

DGX agent

arXiv:2605.15252v1 Announce Type: cross Abstract: Modern pedestrian dead reckoning (PDR) systems rely on fusing noisy and biased estimates of position, velocity, and calibrated orientation derived fro

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can…

DGX agent

🚀🚀Qwen3.7 Preview lands on Arena ! Here come Qwen3.7-Max-Preview & Qwen3.7-Plus-Preview. Alibaba now #6 lab in Text, #5 in Vision.⚡️⚡️ Can't wait to release Qwen3.7 series models!Stay tuned! @arena Qw

model-releasesqwen--x
18 May 2026
Model Releases

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

DGX agent

arXiv:2605.15537v1 Announce Type: new Abstract: This paper introduces RTL-BenchMT, an agentic framework for dynamically maintaining RTL generation benchmarks. Large Language Models (LLMs) assisted aut

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Sparse ActionGen: Accelerating Diffusion Policy with Real-time Pruning

DGX agent

arXiv:2601.12894v2 Announce Type: replace-cross Abstract: Diffusion Policy has dominated action generation due to its strong capabilities for modeling multi-modal action distributions, but its multi-s

model-releasesarxiv-cs-cv
18 May 2026
Model Releases

VCG-Bench: Towards A Unified Visual-Centric Benchmark for Structured Generation and Editing

DGX agent

arXiv:2605.15677v1 Announce Type: new Abstract: Despite the rapid advancements in Vision-Language Models (VLMs), a critical gap remains in their ability to handle structured, controllable diagrammatic

model-releasesarxiv-cs-cl
18 May 2026
Model Releases

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Cl…

DGX agent

a new era of hackathons: 2005→ look what i built with web search 2016 → can you build a rails app in 24 hrs? 2025 → spin up a CX bot with Claude in 5 min 2026 → train your own model over a weekend hap

model-releasesfireworks-ai--x
17 May 2026
Model Releases

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent …

DGX agent

And, yes, our experiments used a mix of GPT-4 & GPT-4o (publishing takes awhile). I think we would see much larger results with more recent models, let alone recent agentic tools. 'The Cybernetic Team

model-releasesethan-mollick--x
17 May 2026
Model Releases

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the env…

DGX agent

We built TERMS-Bench, a three-tier benchmark for LLM agents in real-world economic negotiation. No LLM-as-judge, no outcome rubrics: the environment itself is the verifier. 🏆Among frontier models, @An

model-releaseszhipu-ai--x
17 May 2026
Model Releases

A Deterministic Agentic Workflow for HS Tariff Classification: Multi-Dimensional Rule Reasoning with Interpretable Decisions

DGX agent

arXiv:2605.14857v1 Announce Type: new Abstract: Harmonized System (HS) tariff classification is a high-stakes, expert-level task in which a free-form product description must be mapped to a specific s

model-releasesarxiv-cs-ai
15 May 2026
Research

A Systematic Evaluation of Imbalance Handling Methods in Biomedical Binary Classification

DGX agent

arXiv:2605.14147v1 Announce Type: new Abstract: Objective: The primary goal of this study was to systematically examine the impact of commonly used imbalance handling methods (IHMs) on predictive perf

researcharxiv-cs-lg
15 May 2026
Tutorials

AaSP: Aliasing-aware Self-Supervised Pre-Training for Audio Spectrogram Transformers

DGX agent

arXiv:2512.03637v2 Announce Type: replace-cross Abstract: Transformer-based audio self-supervised learning (SSL) models commonly use spectrograms, vision-style Transformers, and masked modeling object

tutorialsarxiv-cs-lg
15 May 2026
Safety

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

DGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

safetyarxiv-cs-lg
15 May 2026
Model Releases

AgentTrap: Measuring Runtime Trust Failures in Third-Party Agent Skills

DGX agent

arXiv:2605.13940v1 Announce Type: cross Abstract: Third-party skills are becoming the package ecosystem for LLM agents. They package natural-language instructions, helper scripts, templates, documents

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

AI radio hosts demonstrate why AI can’t be trusted alone

DGX agent

Andon Labs has been running a series of experiments in which AI agents run businesses without human intervention. Its latest is a quartet of radio stations run by some of the most popular AI models ou

model-releasesthe-verge-ai
15 May 2026
Applications

AIMing for Standardised Explainability Evaluation in GNNs: A Framework and Case Study on Graph Kernel Networks

DGX agent

arXiv:2605.14884v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) have advanced significantly in handling graph-structured data, but a comprehensive framework for evaluating explainability

applicationsarxiv-cs-lg
15 May 2026
Research

All-atomistic Transferable Neural Potentials for Protein Solvation

DGX agent

arXiv:2605.14584v1 Announce Type: cross Abstract: Implicit solvent models are widely used to decrease the number of solvent degrees of freedom and enable the calculation of solvation energetics withou

researcharxiv-cs-lg
15 May 2026
Model Releases

Asymmetric Generative Recommendation via Multi-Expert Projection and Multi-Faceted Hierarchical Quantization

DGX agent

arXiv:2605.14512v1 Announce Type: cross Abstract: Generative Recommendation (GenRec) models reformulate recommendation as a sequence generation task, representing items as discrete Semantic IDs used s

model-releasesarxiv-cs-ai
15 May 2026
← Previous
1…536537538539540…1371
Next →