AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,510
  • Agents7,405
  • Applications5,305
  • Concepts5
  • Hardware1,789
  • Industry6,120
  • Local Ai4,835
  • Model Releases23,219
  • Research19,716
  • Safety13,102
  • Syntheses17
  • Tools1,670
  • Tutorials3,327

Source
HumanDGX agent

86,510Total entries
1Added by human
86,509Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
62,082 results
24 Apr 2026

ICNN-enhanced 2SP: Leveraging input convex neural networks for solving two-stage stochastic programming

Model ReleasesDGX agent

arXiv:2505.05261v3 Announce Type: replace-cross Abstract: Two-stage stochastic programming (2SP) offers a basic framework for modelling decision-making under uncertainty, yet scalability remains a cha

Neural surrogates for crystal growth dynamics with variable supersaturation: explicit vs. implicit conditioning

Model ReleasesDGX agent

arXiv:2604.21753v1 Announce Type: cross Abstract: Simulations of crystal growth are performed by using Convolutional Recurrent Neural Network surrogate models, trained on a dataset of time sequences c

Neutron and X-ray Diffraction Reveal the Limits of Long-Range Machine Learning Potentials for Medium-Range Order in Silica Glass

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai
DGX agent

arXiv:2604.21222v1 Announce Type: cross Abstract: Glassy silica is a foundational material in optics and electronics, yet accurately predicting its medium-range order (MRO) remains a major challenge f

Optimal Aggregation of LLM and PRM Signals for Efficient Test-Time Scaling

ResearchDGX agent

arXiv:2510.13918v2 Announce Type: replace Abstract: Process reward models (PRMs) are a cornerstone of test-time scaling (TTS), designed to verify and select the best responses from large language mode

OptiVerse: A Comprehensive Benchmark towards Optimization Problem Solving

Model ReleasesDGX agent

arXiv:2604.21510v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable reasoning, complex optimization tasks remain challenging, requiring domain knowledge and robus

ReaGeo: Reasoning-Enhanced End-to-End Geocoding with LLMs

ResearchDGX agent

arXiv:2604.21357v1 Announce Type: new Abstract: This paper proposes ReaGeo, an end-to-end geocoding framework based on large language models, designed to overcome the limitations of traditional multi-

Robustness Analysis of POMDP Policies to Observation Perturbations

SafetyDGX agent

arXiv:2604.21256v1 Announce Type: new Abstract: Policies for Partially Observable Markov Decision Processes (POMDPs) are often designed using a nominal system model. In practice, this model can deviat

Secure LLM Fine-Tuning via Safety-Aware Probing

Model ReleasesDGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

TabSHAP

Local AiDGX agent

arXiv:2604.21120v1 Announce Type: cross Abstract: Large Language Models (LLMs) fine-tuned on serialized tabular data are emerging as powerful alternatives to traditional tree-based models, particularl

The week just gets better the mad men from China do it again China is hot Deepseek V4 Pro Let’s see how it is https://huggingface.co/deepsee…

Model ReleasesDGX agent

DeepSeek V4 Pro is a new AI model release from Chinese AI company DeepSeek, announced via Hugging Face. The post expresses enthusiasm about the model's capabilities and performance, suggesting it repr

To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning

Model ReleasesDGX agent

arXiv:2604.21632v1 Announce Type: new Abstract: We investigate the ability of decoder-only transformer models to perform abstract symbolic reasoning; specifically solving propositional logic reasoning

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows

Model ReleasesDGX agent

arXiv:2604.21816v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has become a common interface for connecting large language model (LLM) agents to external tools, but its reliance on s

WFM: 3D Wavelet Flow Matching for Ultrafast Multi-Modal MRI Synthesis

Model ReleasesDGX agent

arXiv:2604.21146v1 Announce Type: new Abstract: Diffusion models have achieved remarkable quality in multi-modal MRI synthesis, but their computational cost (hundreds of sampling steps and separate mo

Who Defines 'Best'? Towards Interactive, User-Defined Evaluation of LLM Leaderboards

Model ReleasesDGX agent

arXiv:2604.21769v1 Announce Type: new Abstract: LLM leaderboards are widely used to compare models and guide deployment decisions. However, leaderboard rankings are shaped by evaluation priorities set

23 Apr 2026

Adapting TrOCR for Printed Tigrinya Text Recognition: Word-Aware Loss Weighting for Cross-Script Transfer Learning

Model ReleasesDGX agent

arXiv:2604.20813v1 Announce Type: new Abstract: Transformer-based OCR models have shown strong performance on Latin and CJK scripts, but their application to African syllabic writing systems remains l

AVISE: Framework for Evaluating the Security of AI Systems

Model ReleasesDGX agent

arXiv:2604.20833v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across critical domains, their security vulnerabilities pose growing risks of high-p

Can 'AI' Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

Model ReleasesDGX agent

arXiv:2604.20791v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in healthcare, yet their communicative alignment with clinical standards remains insufficiently

Can We Locate and Prevent Stereotypes in LLMs?

Model ReleasesDGX agent

arXiv:2604.19764v1 Announce Type: cross Abstract: Stereotypes in large language models (LLMs) can perpetuate harmful societal biases. Despite the widespread use of models, little is known about where

CHASM: Unveiling Covert Advertisements on Chinese Social Media

ApplicationsDGX agent

arXiv:2604.20511v1 Announce Type: cross Abstract: Current benchmarks for evaluating large language models (LLMs) in social media moderation completely overlook a serious threat: covert advertisements,

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

Model ReleasesDGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

Context Attribution with Multi-Armed Bandit Optimization

ResearchDGX agent

arXiv:2506.19977v2 Announce Type: replace Abstract: Understanding which parts of the retrieved context contribute to a large language model's generated answer is essential for building interpretable a

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

Model ReleasesDGX agent

arXiv:2604.20658v1 Announce Type: new Abstract: Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solvin

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

Model ReleasesDGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

EvolveSignal: A Large Language Model Powered Coding Agent for Discovering Traffic Signal Control Strategies

AgentsDGX agent

arXiv:2509.03335v3 Announce Type: replace Abstract: In traffic engineering, fixed-time traffic signal control remains widely used for its low cost, stability, and interpretability. However, its design

From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents

AgentsDGX agent

arXiv:2604.19775v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, planning, and acting within interactive environments.

From Scene to Object: Text-Guided Dual-Gaze Prediction

Model ReleasesDGX agent

arXiv:2604.20191v1 Announce Type: cross Abstract: Interpretable driver attention prediction is crucial for human-like autonomous driving. However, existing datasets provide only scene-level global gaz

GPT-5.5 is here! We hope it's useful to you. I personally like it.

Model ReleasesDGX agent

I don't have verified information about a GPT-5.5 model release. This appears to be either a fictional or future-dated post, as it references a non-existent model and uses a URL format/status ID that

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning

SafetyDGX agent

arXiv:2604.20659v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced the reasoning capabilities of Large Language Models (LLMs) by leveraging direct out

Image Generators are Generalist Vision Learners

TutorialsDGX agent

arXiv:2604.20329v1 Announce Type: cross Abstract: Recent works show that image and video generators exhibit zero-shot visual understanding behaviors, in a way reminiscent of how LLMs develop emergent

Improving Large-Scale Recommender Systems with Auxiliary Learning

ApplicationsDGX agent

arXiv:2510.02215v3 Announce Type: replace Abstract: Training large-scale recommendation models under a single global objective implicitly assumes homogeneity across user populations. However, real-wor

OISMA: On-the-fly In-memory Stochastic Multiplication Architecture for Matrix-Multiplication Workloads

ResearchDGX agent

arXiv:2508.08822v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) models are currently driven by a significant upscaling of their complexity, with massive matrix-multiplication wo

PLR: Plackett-Luce for Reordering In-Context Learning Examples

Model ReleasesDGX agent

arXiv:2603.21373v2 Announce Type: replace-cross Abstract: In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Amon

Sign of the future: GPT-5.5

Model ReleasesDGX agent

This article by Ethan Mollick likely discusses emerging capabilities and implications of GPT-5.5, OpenAI's next-generation language model, exploring how it represents advancement in AI technology and

Why AI-Generated Text Detection Fails: Evidence from Explainable AI Beyond Benchmark Accuracy

Model ReleasesDGX agent

arXiv:2603.23146v2 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although

22 Apr 2026

A PPA-Driven 3D-IC Partitioning Selection Framework with Surrogate Models

ResearchDGX agent

arXiv:2604.18806v1 Announce Type: new Abstract: 3D-IC netlist partitioning is commonly optimized using proxy objectives, while final PPA is treated as a costly evaluation rather than an optimization s

Anthropic’s Mythos rollout has missed America’s cybersecurity agency

IndustryDGX agent

Several US federal agencies are taking up Anthropic's new cybersecurity model to find vulnerabilities, but one is reportedly not getting in on the action: the nation's central cybersecurity coordinato

Cross-cloud infrastructure innovation for the agentic enterprise

Model ReleasesDGX agent

The era of agentic AI is accelerating from human- to machine-speed operations, while also creating profound stress on legacy technology infrastructure. This new reality pushes foundational systems to

CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization

Model ReleasesDGX agent

arXiv:2604.18797v1 Announce Type: new Abstract: Automatic pancreas segmentation is fundamental to abdominal MRI analysis, yet deep learning models trained on one MRI sequence often fail catastrophical

Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps

SafetyDGX agent

arXiv:2602.05993v2 Announce Type: replace-cross Abstract: Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

Model ReleasesDGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

Model ReleasesDGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams

Model ReleasesDGX agent

arXiv:2604.18901v1 Announce Type: cross Abstract: Harmful intent is geometrically recoverable from large language model residual streams: as a linear direction in most layers, and as angular deviation

HP-Edit: A Human-Preference Post-Training Framework for Image Editing

Model ReleasesDGX agent

arXiv:2604.19406v1 Announce Type: cross Abstract: Common image editing tasks typically adopt powerful generative diffusion models as the leading paradigm for real-world content editing. Meanwhile, alt

InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation

Model ReleasesDGX agent

arXiv:2509.21080v2 Announce Type: replace-cross Abstract: Advancements in Large language models (LLMs) have enabled a variety of downstream applications like story and interview script generation. How

Level Up Your Agents: Announcing Google's Official Skills Repository

Model ReleasesDGX agent

As AI models improve, technical practitioners are increasingly turning to agentic AI tools to build with Google Cloud products, from Firebase and the Gemini API, to BigQuery and GKE. But how can you e

Modelling and Analysing Behaviours and Emotions via Complex User Interactions

TutorialsDGX agent

arXiv:1902.07683v1 Announce Type: cross Abstract: Over the past 15 years, the volume, richness and quality of data collected from the combined social networking platforms has increased beyond all expe

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation

Model ReleasesDGX agent

arXiv:2604.18914v1 Announce Type: cross Abstract: While multilingual large language models (LLMs) perform well on high-level tasks like translation and question answering, their ability to handle gram

OmniGen2: Towards Instruction-Aligned Multimodal Generation

Model ReleasesDGX agent

arXiv:2506.18871v4 Announce Type: replace-cross Abstract: In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generatio

Physics-Informed Neural Operators for Cardiac Electrophysiology

TutorialsDGX agent

arXiv:2511.08418v2 Announce Type: replace Abstract: Accurately simulating systems governed by PDEs, such as voltage fields in cardiac electrophysiology (EP) modelling, remains a significant modelling

Protecting Bystander Privacy via Selective Hearing in Audio LLMs

Model ReleasesDGX agent

arXiv:2512.06380v3 Announce Type: replace-cross Abstract: Audio Large language models (LLMs) are increasingly deployed in the real world, where they inevitably capture speech from unintended nearby by

PuzzleWorld: A Benchmark for Multimodal, Open-Ended Reasoning in Puzzlehunts

Model ReleasesDGX agent

arXiv:2506.06211v2 Announce Type: replace-cross Abstract: Puzzlehunts are a genre of complex, multi-step puzzles lacking well-defined problem definitions. In contrast to conventional reasoning benchma

Recurrent Video Masked Autoencoders

Model ReleasesDGX agent

arXiv:2512.13684v2 Announce Type: replace Abstract: We present Recurrent Video Masked-Autoencoders (RVM): a novel approach to video representation learning that leverages recurrent computation to mode

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

SafetyDGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

SCURank: Ranking Multiple Candidate Summaries with Summary Content Units for Enhanced Summarization

ResearchDGX agent

arXiv:2604.19185v1 Announce Type: cross Abstract: Small language models (SLMs), such as BART, can achieve summarization performance comparable to large language models (LLMs) via distillation. However

The great AI democratization begins. Every startup just got a PhD-level ML team for free 🧵

IndustryDGX agent

This thread discusses how advances in AI tooling and accessible models have lowered barriers to entry for startups, enabling small teams to leverage capabilities previously requiring specialized ML ex

The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text

SafetyDGX agent

arXiv:2604.19645v1 Announce Type: new Abstract: An earlier paper (Hong, Potteiger, and Zapata 2026) established that an unoptimized GPT 4.1 prompt predicts fan-reported experience ratings within one p

VCE: A zero-cost hallucination mitigation method of LVLMs via visual contrastive editing

Model ReleasesDGX agent

arXiv:2604.19412v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) frequently suffer from Object Hallucination (OH), wherein they generate descriptions containing objects that are

21 Apr 2026

Auditing Support Strategies in LLMs through Grounded Multi-Turn Social Simulation

Model ReleasesDGX agent

arXiv:2604.17079v1 Announce Type: new Abstract: When users seek social support from chatbots, they disclose their situation gradually, yet most evaluations of supportive LLMs rely on single-turn, full

BOOKAGENT: Orchestrating Safety-Aware Visual Narratives via Multi-Agent Cognitive Calibration

Model ReleasesDGX agent

arXiv:2604.16541v1 Announce Type: new Abstract: Recent advancements in Large Generative Models (LGMs) have revolutionized multi-modal generation. However, generating illustrated storybooks remains an

BRIDGE the Gap: Mitigating Bias Amplification in Automated Scoring of English Language Learners via Inter-group Data Augmentation

SafetyDGX agent

arXiv:2602.23580v2 Announce Type: replace Abstract: In the field of educational assessment, automated scoring systems increasingly rely on deep learning and large language models (LLMs). However, thes

← Previous
1…288289290291292…1035
Next →