AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries92,413
  • Agents7,866
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,240
  • Local Ai5,175
  • Model Releases25,275
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,515

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries92,413
  • Agents7,866
  • Applications5,605
  • Concepts5
  • Hardware1,963
  • Industry6,240
  • Local Ai5,175
  • Model Releases25,275
  • Research21,121
  • Safety13,951
  • Syntheses17
  • Tools1,680
  • Tutorials3,515

Source
HumanDGX agent

Content type
92,413Total entries
1Added by human
92,412Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
66,933 results
Research

PanoSAMic: Panoramic Image Segmentation from SAM Feature Encoding and Dual View Fusion

DGX agent

arXiv:2601.07447v2 Announce Type: replace Abstract: Existing image foundation models are not optimized for spherical images having been trained primarily on perspective images. PanoSAMic integrates th

researcharxiv-cs-cv
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

DGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

model-releasesarxiv-cs-ai
14 Apr 2026
Research

PAS: Estimating the target accuracy before domain adaptation

DGX agent

arXiv:2604.09863v1 Announce Type: cross Abstract: The goal of domain adaptation is to make predictions for unlabeled samples from a target domain with the help of labeled samples from a different but

researcharxiv-cs-ai
14 Apr 2026
Research

Poisoning with A Pill: Circumventing Detection in Federated Learning

DGX agent

arXiv:2407.15389v2 Announce Type: replace Abstract: Without direct access to the client's data, federated learning (FL) is well-known for its unique strength in data privacy protection among existing

researcharxiv-cs-lg
14 Apr 2026
Local Ai

Post-Processing Methods for Improving Accuracy in MRI Inpainting

DGX agent

arXiv:2510.15282v2 Announce Type: replace-cross Abstract: Magnetic Resonance Imaging (MRI) is the primary imaging modality used in the diagnosis, assessment, and treatment planning for brain pathologi

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

Precision Synthesis of Multi-Tracer PET via VLM-Modulated Rectified Flow for Stratifying Mild Cognitive Impairment

DGX agent

arXiv:2604.11176v1 Announce Type: new Abstract: The biological definition of Alzheimer's disease (AD) relies on multi-modal neuroimaging, yet the clinical utility of positron emission tomography (PET)

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

PRIX: Learning to Plan from Raw Pixels for End-to-End Autonomous Driving

DGX agent

arXiv:2507.17596v3 Announce Type: replace-cross Abstract: While end-to-end autonomous driving models show promising results, their practical deployment is often hindered by large model sizes, a relian

agentsarxiv-cs-ai
14 Apr 2026
Tutorials

Probabilistic Prediction of Neural Dynamics via Autoregressive Flow Matching

DGX agent

arXiv:2604.11178v1 Announce Type: cross Abstract: Forecasting neural activity in response to naturalistic stimuli remains a key challenge for understanding brain dynamics and enabling downstream neuro

tutorialsarxiv-cs-lg
14 Apr 2026
Local Ai

Psychological Concept Neurons: Can Neural Control Bias Probing and Shift Generation in LLMs?

DGX agent

arXiv:2604.11802v1 Announce Type: new Abstract: Using psychological constructs such as the Big Five, large language models (LLMs) can imitate specific personality profiles and predict a user's persona

local-aiarxiv-cs-cl
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

RealSR-R1: Reinforcement Learning for Real-World Image Super-Resolution with Vision-Language Chain-of-Thought

DGX agent

arXiv:2506.16796v4 Announce Type: replace Abstract: Real-World Image Super-Resolution is one of the most challenging task in image restoration. However, existing methods struggle with an accurate unde

safetyarxiv-cs-cv
14 Apr 2026
Model Releases

Reasoning as Gradient: Scaling MLE Agents Beyond Tree Search

DGX agent

arXiv:2603.01692v3 Announce Type: replace-cross Abstract: LLM-based agents for machine learning engineering (MLE) predominantly rely on tree search, a form of gradient-free optimization that uses scal

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ReFEree: Reference-Free and Fine-Grained Method for Evaluating Factual Consistency in Real-World Code Summarization

DGX agent

arXiv:2604.10520v1 Announce Type: cross Abstract: As Large Language Models (LLMs) have become capable of generating long and descriptive code summaries, accurate and reliable evaluation of factual con

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Relax: An Asynchronous Reinforcement Learning Engine for Omni-Modal Post-Training at Scale

DGX agent

arXiv:2604.11554v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has proven effective at unlocking reasoning, self-reflection, and tool-use capabilities in large language mode

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

ReXSonoVQA: A Video QA Benchmark for Procedure-Centric Ultrasound Understanding

DGX agent

arXiv:2604.10916v1 Announce Type: cross Abstract: Ultrasound acquisition requires skilled probe manipulation and real-time adjustments. Vision-language models (VLMs) could enable autonomous ultrasound

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

RL makes MLLMs see better than SFT

DGX agent

arXiv:2510.16333v2 Announce Type: replace Abstract: A dominant assumption in Multimodal Language Model (MLLM) research is that its performance is largely inherited from the LLM backbone, given its imm

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Robust Fair Disease Diagnosis in CT Images

DGX agent

arXiv:2604.09710v1 Announce Type: new Abstract: Automated diagnosis from chest CT has improved considerably with deep learning, but models trained on skewed datasets tend to perform unevenly across pa

model-releasesarxiv-cs-cv
14 Apr 2026
Local Ai

RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents

DGX agent

arXiv:2604.11655v1 Announce Type: cross Abstract: The rapid adoption of Large Language Models (LLMs) in interactive systems has enabled the creation of dynamic, open-ended Role-Playing Agents (RPAs).

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SecureVibeBench: Evaluating Secure Coding Capabilities of Code Agents with Realistic Vulnerability Scenarios

DGX agent

arXiv:2509.22097v3 Announce Type: replace-cross Abstract: Large language model-powered code agents are rapidly transforming software engineering, yet the security risks of their generated code have be

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Semantic-Geometric Dual Compression: Training-Free Visual Token Reduction for Ultra-High-Resolution Remote Sensing Understanding

DGX agent

arXiv:2604.11122v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have demonstrated immense potential in Earth observation. However, the massive visual tokens generated when p

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Sense Less, Infer More: Agentic Multimodal Transformers for Edge Medical Intelligence

DGX agent

arXiv:2604.10404v1 Announce Type: cross Abstract: Edge-based multimodal medical monitoring requires models that balance diagnostic accuracy with severe energy constraints. Continuous acquisition of EC

safetyarxiv-cs-lg
14 Apr 2026
Model Releases

Solving Physics Olympiad via Reinforcement Learning on Physics Simulators

DGX agent

arXiv:2604.11805v1 Announce Type: cross Abstract: We have witnessed remarkable advances in LLM reasoning capabilities with the advent of DeepSeek-R1. However, much of this progress has been fueled by

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Steered LLM Activations are Non-Surjective

DGX agent

arXiv:2604.09839v1 Announce Type: new Abstract: Activation steering is a popular white-box control technique that modifies model activations to elicit an abstract change in output behavior. It has als

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Switch-JustDance: Benchmarking Whole Body Motion Tracking Controllers Using a Commercial Console Game

DGX agent

arXiv:2511.17925v3 Announce Type: replace-cross Abstract: Recent advances in whole-body robot control have enabled humanoid and legged robots to perform increasingly agile and coordinated motions. How

model-releasesarxiv-cs-cv
14 Apr 2026
Research

TARAC: Mitigating Hallucination in LVLMs via Temporal Attention Real-time Accumulative Connection

DGX agent

arXiv:2504.04099v2 Announce Type: replace-cross Abstract: Large Vision-Language Models have demonstrated remarkable capabilities, yet they suffer from hallucinations that limit practical deployment. W

researcharxiv-cs-ai
14 Apr 2026
Research

Teaching Robots to Interpret Social Interactions through Lexically-guided Dynamic Graph Learning

DGX agent

arXiv:2604.10895v1 Announce Type: cross Abstract: For a robot to be called socially intelligent, it must be able to infer users internal states from their current behaviour, predict the users future b

researcharxiv-cs-ro
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Think Parallax: Solving Multi-Hop Problems via Multi-View Knowledge-Graph-Based Retrieval-Augmented Generation

DGX agent

arXiv:2510.15552v3 Announce Type: replace-cross Abstract: Large language models (LLMs) still struggle with multi-hop reasoning over knowledge-graphs (KGs), and we identify a previously overlooked stru

model-releasesarxiv-cs-ai
14 Apr 2026
Industry

Time to follow http://hf.co/tencent!

DGX agent

Time to follow http://hf.co/tencent! Genie3 generates videos. We generate 𝟯𝗗 𝘄𝗼𝗿𝗹𝗱𝘀 you can actually use. Launching tomorrow — Tencent #HYWorld 2.0, an engine-ready World Model🚀 This isn't a video. It

industryclem-delangue--x
14 Apr 2026
Model Releases

Topo-ADV: Generating Topology-Driven Imperceptible Adversarial Point Clouds

DGX agent

arXiv:2604.09879v1 Announce Type: new Abstract: Deep neural networks for 3D point cloud understanding have achieved remarkable success in object classification and recognition, yet recent work shows t

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Training-Free Object-Background Compositional T2I via Dynamic Spatial Guidance and Multi-Path Pruning

DGX agent

arXiv:2604.09850v1 Announce Type: new Abstract: Existing text-to-image diffusion models, while excelling at subject synthesis, exhibit a persistent foreground bias that treats the background as a pass

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

Trusted access for the next era of cyber defense

DGX agent

OpenAI is expanding trusted access to its AI models for cybersecurity purposes, enabling vetted researchers, defenders, and organizations to leverage advanced AI capabilities for cyber defense applica

model-releasesopenai
14 Apr 2026
Safety

Weird Generalization is Weirdly Brittle

DGX agent

arXiv:2604.10022v1 Announce Type: new Abstract: Weird generalization is a phenomenon in which models fine-tuned on data from a narrow domain (e.g. insecure code) develop surprising traits that manifes

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

What and Where to Adapt: Structure-Semantics Co-Tuning for Machine Vision Compression via Synergistic Adapters

DGX agent

arXiv:2604.10017v1 Announce Type: new Abstract: Parameter-efficient fine-tuning of pre-trained codecs is a promising direction in image compression for human and machine vision. While most existing wo

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

What Factors Affect LLMs and RLLMs in Financial Question Answering?

DGX agent

arXiv:2507.08339v4 Announce Type: replace Abstract: Recently, large language models (LLMs) and reasoning large language models (RLLMs) have gained considerable attention from many researchers. RLLMs e

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Who Wrote This Line? Evaluating the Detection of LLM-Generated Classical Chinese Poetry

DGX agent

arXiv:2604.10101v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has extended text generation tasks into the literary domain. However, AI-generated literary creati

model-releasesarxiv-cs-cl
14 Apr 2026
Research

XD-MAP: Cross-Modal Domain Adaptation via Semantic Parametric Maps for Scalable Training Data Generation

DGX agent

arXiv:2601.14477v2 Announce Type: replace-cross Abstract: Until open-world foundation models match the performance of specialized approaches, deep learning systems remain dependent on task- and sensor

researcharxiv-cs-ai
14 Apr 2026
Model Releases

4D-RGPT: Toward Region-level 4D Understanding via Perceptual Distillation

DGX agent

arXiv:2512.17012v3 Announce Type: replace Abstract: Despite advances in Multimodal LLMs (MLLMs), their ability to reason over 3D structures and temporal dynamics remains limited, constrained by weak 4

model-releasesarxiv-cs-cv
13 Apr 2026
Research

Across the Levels of Analysis: Explaining Predictive Processing in Humans Requires More Than Machine-Estimated Probabilities

DGX agent

arXiv:2604.09466v1 Announce Type: new Abstract: Under the lens of Marr's levels of analysis, we critique and extend two claims about language models (LMs) and language processing: first, that predicti

researcharxiv-cs-cl
13 Apr 2026
Model Releases

AgentCE-Bench: Agent Configurable Evaluation with Scalable Horizons and Controllable Difficulty under Lightweight Environments

DGX agent

arXiv:2604.06111v2 Announce Type: replace Abstract: Existing Agent benchmarks suffer from two critical limitations: high environment interaction overhead (up to 41% of total evaluation time) and imbal

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Automated Instruction Revision (AIR): A Structured Comparison of Task Adaptation Strategies for LLM

DGX agent

arXiv:2604.09418v1 Announce Type: new Abstract: This paper studies Automated Instruction Revision (AIR), a rule-induction-based method for adapting large language models (LLMs) to downstream tasks usi

model-releasesarxiv-cs-cl
13 Apr 2026
Safety

Better Eyes, Better Thoughts: Why Vision Chain-of-Thought Fails in Medicine

DGX agent

arXiv:2603.06665v2 Announce Type: replace-cross Abstract: Large vision-language models (VLMs) often benefit from chain-of-thought (CoT) prompting in general domains, yet its efficacy in medical vision

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

Bharat Scene Text: A Novel Comprehensive Dataset and Benchmark for Indian Language Scene Text Understanding

DGX agent

arXiv:2511.23071v2 Announce Type: replace-cross Abstract: Reading scene text, that is, text appearing in images, has numerous application areas, including assistive technology, search, and e-commerce.

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Boosting Brain-inspired Path Integration Efficiency via Learning-based Replication of Continuous Attractor Neurodynamics

DGX agent

arXiv:2511.17687v2 Announce Type: replace Abstract: The brain's Path Integration (PI) mechanism offers substantial guidance and inspiration for Brain-Inspired Navigation (BIN). However, the PI capabil

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

CAD 100K: A Comprehensive Multi-Task Dataset for Car Related Visual Anomaly Detection

DGX agent

arXiv:2604.09023v1 Announce Type: new Abstract: Multi-task visual anomaly detection is critical for car-related manufacturing quality assessment. However, existing methods remain task-specific, hinder

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

DeFakeQ: Enabling Real-Time Deepfake Detection on Edge Devices via Adaptive Bidirectional Quantization

DGX agent

arXiv:2604.08847v1 Announce Type: new Abstract: Deepfake detection has become a fundamental component of modern media forensics. Despite significant progress in detection accuracy, most existing metho

model-releasesarxiv-cs-cv
13 Apr 2026
← Previous
1…571572573574575…1395
Next →