AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “research”

GridTimelineEvolution
25,428 results
Model Releases

Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads

DGX agent

arXiv:2604.23009v1 Announce Type: new Abstract: Job Skill Named Entity Recognition (JobSkillNER) aims to automatically extract key skill information from large-scale job posting data, which is importa

model-releasesarxiv-cs-cl
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ClawTrace: Cost-Aware Tracing for LLM Agent Skill Distillation

DGX agent

arXiv:2604.23853v1 Announce Type: new Abstract: Skill-distillation pipelines learn reusable rules from LLM agent trajectories, but they lack a key signal: how much each step costs. Without per-step co

model-releasesarxiv-cs-ai
28 Apr 2026
Applications

Comparative Insights on Adversarial Machine Learning from Industry and Academia: A User-Study Approach

DGX agent

arXiv:2602.04753v2 Announce Type: replace-cross Abstract: An exponential growth of Machine Learning and its Generative AI applications brings with it significant security challenges, often referred to

applicationsarxiv-cs-ai
28 Apr 2026
Model Releases

Complexity of Linear Regions in Self-supervised Deep ReLU Networks

DGX agent

arXiv:2604.24393v1 Announce Type: cross Abstract: There has been growing interest in studying the complexity of Rectified Linear Unit (ReLU) based activation networks. Recent work investigates the evo

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Computational Design and Co-Robotic Fabrication for Material Reuse in Architecture

DGX agent

arXiv:2604.24648v1 Announce Type: new Abstract: Climate change and resource depletion demand a shift from the dominant linear 'take-make-use-dispose' paradigm of construction toward circular, low-wast

applicationsarxiv-cs-ro
28 Apr 2026
Model Releases

DGHMesh: A Large-scale Dual-radar mmWave Dataset and Generalization-focused Benchmark for Human Mesh Reconstruction

DGX agent

arXiv:2604.22827v1 Announce Type: new Abstract: Millimeter-wave (mmWave) radar has shown great potential for contactless, privacy-preserving, and robust human sensing, yet existing mmWave-based human

model-releasesarxiv-cs-cv
28 Apr 2026
Applications

Do Protective Perturbations Really Protect Portrait Privacy under Real-world Image Transformations?

DGX agent

arXiv:2604.23688v1 Announce Type: new Abstract: Proactive defense methods protect portrait images from unauthorized editing or talking face generation (TFG) by introducing pixel-level protective pertu

applicationsarxiv-cs-cv
28 Apr 2026
Model Releases

DyABD: The Abdominal Muscle Segmentation in Dynamic MRI Benchmark

DGX agent

arXiv:2604.23187v1 Announce Type: cross Abstract: This work introduces DyABD, a novel and complex benchmark dataset of dynamic abdominal MRIs from patients with abdominal hernias and associated high q

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

DGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

EmoBench-M: Benchmarking Emotional Intelligence for Multimodal Large Language Models

DGX agent

arXiv:2502.04424v4 Announce Type: replace-cross Abstract: With the integration of multimodal large language models (MLLMs) into robotic systems and AI applications, embedding emotional intelligence (E

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

EmoTrans: A Benchmark for Understanding, Reasoning, and Predicting Emotion Transitions in Multimodal LLMs

DGX agent

arXiv:2604.23348v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have shown strong capabilities in perception, reasoning, and generation, and are increasingly used in

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Evaluating the Search Agent in a Parallel World

DGX agent

arXiv:2603.04751v2 Announce Type: replace Abstract: Integrating web search tools has significantly extended the capability of LLMs to address open-world, real-time, and long-tail problems. However, ev

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Exploring the Secondary Risks of Large Language Models

DGX agent

arXiv:2506.12382v5 Announce Type: replace-cross Abstract: Ensuring the safety and alignment of Large Language Models is a significant challenge with their growing integration into critical application

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment

DGX agent

arXiv:2604.23786v1 Announce Type: new Abstract: In recent years, the integration of multimodal machine learning in wellbeing assessment has offered transformative potential for monitoring mental healt

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

FastAT Benchmark: A Comprehensive Framework for Fair Evaluation of Fast Adversarial Training Methods

DGX agent

arXiv:2604.22853v1 Announce Type: new Abstract: Fast Adversarial Training (FastAT) seeks to achieve adversarial robustness at a fraction of the computational cost incurred by standard multi-step metho

model-releasesarxiv-cs-cv
28 Apr 2026
Safety

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

DGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

Forecasting Commencing Enrolments Under Data Sparsity: A Zero-Shot Time Series Foundation Models Framework for Higher Education Planning

DGX agent

arXiv:2602.12120v3 Announce Type: replace Abstract: Effective resource allocation in higher education depends on reliable enrolment forecasts, yet institutional planners frequently face data series di

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

// From Skill Text to Skill Structure // One of the more practical skill papers I've seen this month. SKILL.md files entangle invocation int…

DGX agent

// From Skill Text to Skill Structure // One of the more practical skill papers I've seen this month. SKILL.md files entangle invocation interface, execution flow, and tool/resource side effects in on

tutorialsdair-ai--x
28 Apr 2026
Safety

From Stateless Queries to Autonomous Actions: A Layered Security Framework for Agentic AI Systems

DGX agent

arXiv:2604.23338v1 Announce Type: cross Abstract: Agentic AI systems face security challenges that stateless large language models do not. They plan across extended horizons, maintain persistent memor

safetyarxiv-cs-lg
28 Apr 2026
Model Releases

Game-Time: Evaluating Temporal Dynamics in Spoken Language Models

DGX agent

arXiv:2509.26388v3 Announce Type: replace-cross Abstract: Conversational Spoken Language Models (SLMs) are emerging as a promising paradigm for real-time speech interaction. However, their capacity of

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Google AI 인프라의 미래: 에이전틱 시대를 위한 확장

DGX agent

* 본 아티클의 원문은 2026년 4월 23일 Google Cloud 블로그(영문)에 게재되었습니다. AI는 질문에 답하는 수준을 넘어 추론하고 행동하는 단계로 진화하고 있습니다. 오늘날의 에이전틱 시대(agentic era)를 선도하고자 하는 기업에는 이러한 새로운 요구사항에 맞춰 설계되고 최적화된 컴퓨팅 인프라가 필요합니다. 오늘 Google Cloud

model-releasesgoogle-cloud-ai
28 Apr 2026
Agents

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations -…

DGX agent

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations - curating datasets - running experiments - and iterating lik

agentsclem-delangue--x
28 Apr 2026
Tutorials

'If You're Very Clever, No One Knows You've Used It': The Social Dynamics of Developing Generative AI Literacy in the Workplace

DGX agent

arXiv:2602.01386v2 Announce Type: replace-cross Abstract: Generative AI (GenAI) tools are rapidly transforming knowledge work, making AI literacy a critical priority for organizations. However, resear

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

Introducing talkie: a 13B vintage language model from 1930

DGX agent

Introducing talkie: a 13B vintage language model from 1930 New project from Nick Levine, David Duvenaud, and Alec Radford (of GPT, GPT-2, Whisper fame). talkie-1930-13b-base (53.1 GB) is a '13B langua

model-releasessimon-willison
28 Apr 2026
Model Releases

I've been working on a side project for the last few weeks... what if you could have a Gemma powered app that would let you have a personal …

DGX agent

I've been working on a side project for the last few weeks... what if you could have a Gemma powered app that would let you have a personal assistant that could browse the internet with you, do resear

model-releasesollama--x
28 Apr 2026
Model Releases

Judging the Judges: A Systematic Evaluation of Bias Mitigation Strategies in LLM-as-a-Judge Pipelines

DGX agent

arXiv:2604.23178v1 Announce Type: new Abstract: LLM-as-a-Judge has become the dominant paradigm for evaluating language model outputs, yet LLM judges exhibit systematic biases that compromise evaluati

model-releasesarxiv-cs-ai
28 Apr 2026
Hardware

Latent Inter-Frame Pruning: A Training-Free Method Bridging Traditional Video Compression and Modern Diffusion Transformers for Efficient Generation

DGX agent

arXiv:2604.23858v1 Announce Type: new Abstract: Video generation, while capable of generating realistic videos, is computationally expensive and slow, prohibiting real-time applications. In this paper

hardwarearxiv-cs-cv
28 Apr 2026
Safety

Learning from Imperfect Text Guidance: Robust Long-Tail Visual Recognition with High-Noise Label

DGX agent

arXiv:2604.23125v1 Announce Type: new Abstract: Real-world data often exhibit long-tailed distributions with numerous noisy labels, substantially degrading the performance of deep models. While prior

safetyarxiv-cs-cv
28 Apr 2026
Model Releases

Learning to Conceal Risk: Controllable Multi-turn Red Teaming for LLMs in the Financial Domain

DGX agent

arXiv:2509.10546v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed in finance, where unsafe behavior can lead to serious regulatory risks. However, most r

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Learning to Identify Out-of-Distribution Objects for 3D LiDAR Anomaly Segmentation

DGX agent

arXiv:2604.23604v1 Announce Type: new Abstract: Understanding the surrounding environment is fundamental in autonomous driving and robotic perception. Distinguishing between known classes and previous

agentsarxiv-cs-cv
28 Apr 2026
Model Releases

LLM4SCREENLIT: Recommendations on Assessing the Performance of Large Language Models for Screening Literature in Systematic Reviews

DGX agent

arXiv:2511.12635v2 Announce Type: replace-cross Abstract: Context: Large language models (LLMs) are increasingly used to screen literature for systematic reviews (SRs), but the standard confusion-matr

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

OLaPh: Optimal Language Phonemizer

DGX agent

arXiv:2509.20086v2 Announce Type: replace Abstract: Phonemization is a critical component in text-to-speech synthesis. Traditional approaches rely on deterministic transformations and lexica, while ne

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

On the Surprising Effectiveness of a Single Global Merging in Decentralized Learning

DGX agent

arXiv:2507.06542v4 Announce Type: replace Abstract: Decentralized learning provides a scalable alternative to parameter-server-based training, yet its performance is often hindered by limited peer-to-

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

OptProver: Bridging Olympiad and Optimization through Continual Training in Formal Theorem Proving

DGX agent

arXiv:2604.23712v1 Announce Type: cross Abstract: Recent advances in formal theorem proving have focused on Olympiad-level mathematics, leaving undergraduate domains largely unexplored. Optimization,

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Orthogonal Representation Learning for Estimating Causal Quantities

DGX agent

arXiv:2502.04274v4 Announce Type: replace Abstract: End-to-end representation learning has become a powerful tool for estimating causal quantities from high-dimensional observational data, but its eff

safetyarxiv-cs-lg
28 Apr 2026
Safety

Our commitment to community safety

DGX agent

OpenAI outlines its commitment to implementing safety measures and responsible practices in the development and deployment of AI systems to protect users and communities. The statement likely covers O

safetyopenai
28 Apr 2026
Safety

Overcoming Copyright Barriers in Corpus Distribution Through Non-Reversible Hashing

DGX agent

arXiv:2604.23412v1 Announce Type: new Abstract: While annotated corpora are crucial in the field of natural language processing (NLP), those containing copyrighted material are difficult to exchange a

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

PivotMerge: Bridging Heterogeneous Multimodal Pre-training via Post-Alignment Model Merging

DGX agent

arXiv:2604.22823v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) rely on multimodal pre-training over diverse data sources, where different datasets often induce complementar

model-releasesarxiv-cs-ai
28 Apr 2026
Tutorials

Probing Visual Planning in Image Editing Models

DGX agent

arXiv:2604.22868v1 Announce Type: cross Abstract: Visual planning represents a crucial facet of human intelligence, especially in tasks that require complex spatial reasoning and navigation. Yet, in m

tutorialsarxiv-cs-ai
28 Apr 2026
Model Releases

QEVA: A Reference-Free Evaluation Metric for Narrative Video Summarization with Multimodal Question Answering

DGX agent

arXiv:2604.24052v1 Announce Type: cross Abstract: Video-to-text summarization remains underexplored in terms of comprehensive evaluation methods. Traditional n-gram overlap-based metrics and recent la

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Quantum Knowledge Graph: Modeling Context-Dependent Triplet Validity

DGX agent

arXiv:2604.23972v1 Announce Type: cross Abstract: Knowledge graphs (KGs) are increasingly used to support large lan guage model (LLM) reasoning, but standard triplet-based KGs treat each relation as g

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Reclaiming Residual Knowledge: A Novel Paradigm to Low-Bit Quantization

DGX agent

arXiv:2408.00923v2 Announce Type: replace-cross Abstract: This paper explores a novel paradigm in low-bit (i.e. 4-bits or lower) quantization, differing from existing state-of-the-art methods, by fram

model-releasesarxiv-cs-ai
28 Apr 2026
Local Ai

Scalable Explainability-as-a-Service (XaaS) for Edge AI Systems

DGX agent

arXiv:2602.04120v2 Announce Type: replace-cross Abstract: Though Explainable AI (XAI) has made significant advancements, its inclusion in edge and IoT systems is typically ad-hoc and inefficient. Most

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ServImage: An Image Generation and Editing Benchmark from Real-world Commercial Imaging Services

DGX agent

arXiv:2604.24023v1 Announce Type: new Abstract: Recent image generation and editing models demonstrate robust adherence to instructions and high visual quality on academic benchmarks. However, their p

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

SFT-then-RL Outperforms Mixed-Policy Methods for LLM Reasoning

DGX agent

arXiv:2604.23747v1 Announce Type: cross Abstract: Recent mixed-policy optimization methods for LLM reasoning that interleave or blend supervised and reinforcement learning signals report improvements

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

ShredBench: Evaluating the Semantic Reasoning Capabilities of Multimodal LLMs in Document Reconstruction

DGX agent

arXiv:2604.23813v1 Announce Type: cross Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable performance in Visually Rich Document Understanding (VRDU) tasks, but their capabili

model-releasesarxiv-cs-cl
28 Apr 2026
Model Releases

SIV-Bench: A Video Benchmark for Social Interaction Understanding and Reasoning

DGX agent

arXiv:2506.05425v2 Announce Type: replace-cross Abstract: Understanding social interaction, which encompasses perceiving numerous and subtle multimodal cues, inferring unobservable mental states and r

model-releasesarxiv-cs-ai
28 Apr 2026
← Previous
1…512513514515516…530
Next →