AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,271
  • Agents7,543
  • Applications5,408
  • Concepts5
  • Hardware1,828
  • Industry6,162
  • Local Ai4,927
  • Model Releases23,818
  • Research20,122
  • Safety13,367
  • Syntheses17
  • Tools1,674
  • Tutorials3,400

Source
HumanDGX agent

Content type
AllBlog
88,271Total entries
1Added by human
88,270Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,519 results
Model Releases

Secure LLM Fine-Tuning via Safety-Aware Probing

DGX agent

arXiv:2505.16737v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have achieved remarkable success across many applications, but their ability to generate harmful content raises s

model-releasesarxiv-cs-ai
24 Apr 2026
Local Ai

TabSHAP

DGX agent

arXiv:2604.21120v1 Announce Type: cross Abstract: Large Language Models (LLMs) fine-tuned on serialized tabular data are emerging as powerful alternatives to traditional tree-based models, particularl

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
local-aiarxiv-cs-cl
24 Apr 2026
Model Releases

The week just gets better the mad men from China do it again China is hot Deepseek V4 Pro Let’s see how it is https://huggingface.co/deepsee…

DGX agent

DeepSeek V4 Pro is a new AI model release from Chinese AI company DeepSeek, announced via Hugging Face. The post expresses enthusiasm about the model's capabilities and performance, suggesting it repr

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

To See the Unseen: on the Generalization Ability of Transformers in Symbolic Reasoning

DGX agent

arXiv:2604.21632v1 Announce Type: new Abstract: We investigate the ability of decoder-only transformer models to perform abstract symbolic reasoning; specifically solving propositional logic reasoning

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows

DGX agent

arXiv:2604.21816v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has become a common interface for connecting large language model (LLM) agents to external tools, but its reliance on s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

WFM: 3D Wavelet Flow Matching for Ultrafast Multi-Modal MRI Synthesis

DGX agent

arXiv:2604.21146v1 Announce Type: new Abstract: Diffusion models have achieved remarkable quality in multi-modal MRI synthesis, but their computational cost (hundreds of sampling steps and separate mo

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Who Defines 'Best'? Towards Interactive, User-Defined Evaluation of LLM Leaderboards

DGX agent

arXiv:2604.21769v1 Announce Type: new Abstract: LLM leaderboards are widely used to compare models and guide deployment decisions. However, leaderboard rankings are shaped by evaluation priorities set

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Adapting TrOCR for Printed Tigrinya Text Recognition: Word-Aware Loss Weighting for Cross-Script Transfer Learning

DGX agent

arXiv:2604.20813v1 Announce Type: new Abstract: Transformer-based OCR models have shown strong performance on Latin and CJK scripts, but their application to African syllabic writing systems remains l

model-releasesarxiv-cs-cv
23 Apr 2026
Model Releases

AVISE: Framework for Evaluating the Security of AI Systems

DGX agent

arXiv:2604.20833v1 Announce Type: cross Abstract: As artificial intelligence (AI) systems are increasingly deployed across critical domains, their security vulnerabilities pose growing risks of high-p

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Can 'AI' Be a Doctor? A Study of Empathy, Readability, and Alignment in Clinical LLMs

DGX agent

arXiv:2604.20791v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed in healthcare, yet their communicative alignment with clinical standards remains insufficiently

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Can We Locate and Prevent Stereotypes in LLMs?

DGX agent

arXiv:2604.19764v1 Announce Type: cross Abstract: Stereotypes in large language models (LLMs) can perpetuate harmful societal biases. Despite the widespread use of models, little is known about where

model-releasesarxiv-cs-ai
23 Apr 2026
Applications

CHASM: Unveiling Covert Advertisements on Chinese Social Media

DGX agent

arXiv:2604.20511v1 Announce Type: cross Abstract: Current benchmarks for evaluating large language models (LLMs) in social media moderation completely overlook a serious threat: covert advertisements,

applicationsarxiv-cs-ai
23 Apr 2026
Model Releases

COMPASS: COntinual Multilingual PEFT with Adaptive Semantic Sampling

DGX agent

arXiv:2604.20720v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit performance disparities across languages, with naive multilingual fine-tuning frequently degrading performa

model-releasesarxiv-cs-ai
23 Apr 2026
Research

Context Attribution with Multi-Armed Bandit Optimization

DGX agent

arXiv:2506.19977v2 Announce Type: replace Abstract: Understanding which parts of the retrieved context contribute to a large language model's generated answer is essential for building interpretable a

researcharxiv-cs-ai
23 Apr 2026
Model Releases

Cooperative Profiles Predict Multi-Agent LLM Team Performance in AI for Science Workflows

DGX agent

arXiv:2604.20658v1 Announce Type: new Abstract: Multi-agent systems built from teams of large language models (LLMs) are increasingly deployed for collaborative scientific reasoning and problem-solvin

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

DialToM: A Theory of Mind Benchmark for Forecasting State-Driven Dialogue Trajectories

DGX agent

arXiv:2604.20443v1 Announce Type: cross Abstract: Large Language Models (LLMs) have been shown to possess Theory of Mind (ToM) abilities. However, it remains unclear whether this stems from robust rea

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

EvolveSignal: A Large Language Model Powered Coding Agent for Discovering Traffic Signal Control Strategies

DGX agent

arXiv:2509.03335v3 Announce Type: replace Abstract: In traffic engineering, fixed-time traffic signal control remains widely used for its low cost, stability, and interpretability. However, its design

agentsarxiv-cs-lg
23 Apr 2026
Agents

From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents

DGX agent

arXiv:2604.19775v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, planning, and acting within interactive environments.

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

From Scene to Object: Text-Guided Dual-Gaze Prediction

DGX agent

arXiv:2604.20191v1 Announce Type: cross Abstract: Interpretable driver attention prediction is crucial for human-like autonomous driving. However, existing datasets provide only scene-level global gaz

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

GPT-5.5 is here! We hope it's useful to you. I personally like it.

DGX agent

I don't have verified information about a GPT-5.5 model release. This appears to be either a fictional or future-dated post, as it references a non-existent model and uses a URL format/status ID that

model-releasessam-altman--x
23 Apr 2026
Safety

GRPO-VPS: Enhancing Group Relative Policy Optimization with Verifiable Process Supervision for Effective Reasoning

DGX agent

arXiv:2604.20659v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has advanced the reasoning capabilities of Large Language Models (LLMs) by leveraging direct out

safetyarxiv-cs-ai
23 Apr 2026
Tutorials

Image Generators are Generalist Vision Learners

DGX agent

arXiv:2604.20329v1 Announce Type: cross Abstract: Recent works show that image and video generators exhibit zero-shot visual understanding behaviors, in a way reminiscent of how LLMs develop emergent

tutorialsarxiv-cs-ai
23 Apr 2026
Applications

Improving Large-Scale Recommender Systems with Auxiliary Learning

DGX agent

arXiv:2510.02215v3 Announce Type: replace Abstract: Training large-scale recommendation models under a single global objective implicitly assumes homogeneity across user populations. However, real-wor

applicationsarxiv-cs-lg
23 Apr 2026
Research

OISMA: On-the-fly In-memory Stochastic Multiplication Architecture for Matrix-Multiplication Workloads

DGX agent

arXiv:2508.08822v2 Announce Type: replace-cross Abstract: Artificial intelligence (AI) models are currently driven by a significant upscaling of their complexity, with massive matrix-multiplication wo

researcharxiv-cs-ai
23 Apr 2026
Model Releases

PLR: Plackett-Luce for Reordering In-Context Learning Examples

DGX agent

arXiv:2603.21373v2 Announce Type: replace-cross Abstract: In-context learning (ICL) adapts large language models by conditioning on a small set of ICL examples, avoiding costly parameter updates. Amon

model-releasesarxiv-cs-cl
23 Apr 2026
Model Releases

Sign of the future: GPT-5.5

DGX agent

This article by Ethan Mollick likely discusses emerging capabilities and implications of GPT-5.5, OpenAI's next-generation language model, exploring how it represents advancement in AI technology and

model-releasesethan-mollick
23 Apr 2026
Model Releases

Why AI-Generated Text Detection Fails: Evidence from Explainable AI Beyond Benchmark Accuracy

DGX agent

arXiv:2603.23146v2 Announce Type: replace-cross Abstract: The widespread adoption of Large Language Models (LLMs) has made the detection of AI-Generated text a pressing and complex challenge. Although

model-releasesarxiv-cs-ai
23 Apr 2026
Research

A PPA-Driven 3D-IC Partitioning Selection Framework with Surrogate Models

DGX agent

arXiv:2604.18806v1 Announce Type: new Abstract: 3D-IC netlist partitioning is commonly optimized using proxy objectives, while final PPA is treated as a costly evaluation rather than an optimization s

researcharxiv-cs-lg
22 Apr 2026
Industry

Anthropic’s Mythos rollout has missed America’s cybersecurity agency

DGX agent

Several US federal agencies are taking up Anthropic's new cybersecurity model to find vulnerabilities, but one is reportedly not getting in on the action: the nation's central cybersecurity coordinato

industrythe-verge-ai
22 Apr 2026
Model Releases

Cross-cloud infrastructure innovation for the agentic enterprise

DGX agent

The era of agentic AI is accelerating from human- to machine-speed operations, while also creating profound stress on legacy technology infrastructure. This new reality pushes foundational systems to

model-releasesgoogle-cloud-ai
22 Apr 2026
Model Releases

CrossPan: A Comprehensive Benchmark for Cross-Sequence Pancreas MRI Segmentation and Generalization

DGX agent

arXiv:2604.18797v1 Announce Type: new Abstract: Automatic pancreas segmentation is fundamental to abdominal MRI analysis, yet deep learning models trained on one MRI sequence often fail catastrophical

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps

DGX agent

arXiv:2602.05993v2 Announce Type: replace-cross Abstract: Flow and diffusion models produce high-quality samples, but adapting them to user preferences or constraints post-training remains costly and

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

DGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HarDBench: A Benchmark for Draft-Based Co-Authoring Jailbreak Attacks for Safe Human-LLM Collaborative Writing

DGX agent

arXiv:2604.19274v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as co-authors in collaborative writing, where users begin with rough drafts and rely on LLMs to compl

model-releasesarxiv-cs-cl
22 Apr 2026
Model Releases

Harmful Intent as a Geometrically Recoverable Feature of LLM Residual Streams

DGX agent

arXiv:2604.18901v1 Announce Type: cross Abstract: Harmful intent is geometrically recoverable from large language model residual streams: as a linear direction in most layers, and as angular deviation

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

HP-Edit: A Human-Preference Post-Training Framework for Image Editing

DGX agent

arXiv:2604.19406v1 Announce Type: cross Abstract: Common image editing tasks typically adopt powerful generative diffusion models as the leading paradigm for real-world content editing. Meanwhile, alt

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

InsideOut: Measuring and Mitigating Insider-Outsider Bias in Interview Script Generation

DGX agent

arXiv:2509.21080v2 Announce Type: replace-cross Abstract: Advancements in Large language models (LLMs) have enabled a variety of downstream applications like story and interview script generation. How

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Level Up Your Agents: Announcing Google's Official Skills Repository

DGX agent

As AI models improve, technical practitioners are increasingly turning to agentic AI tools to build with Google Cloud products, from Firebase and the Gemini API, to BigQuery and GKE. But how can you e

model-releasesgoogle-cloud-ai
22 Apr 2026
Tutorials

Modelling and Analysing Behaviours and Emotions via Complex User Interactions

DGX agent

arXiv:1902.07683v1 Announce Type: cross Abstract: Over the past 15 years, the volume, richness and quality of data collected from the combined social networking platforms has increased beyond all expe

tutorialsarxiv-cs-ai
22 Apr 2026
Model Releases

MORPHOGEN: A Multilingual Benchmark for Evaluating Gender-Aware Morphological Generation

DGX agent

arXiv:2604.18914v1 Announce Type: cross Abstract: While multilingual large language models (LLMs) perform well on high-level tasks like translation and question answering, their ability to handle gram

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

OmniGen2: Towards Instruction-Aligned Multimodal Generation

DGX agent

arXiv:2506.18871v4 Announce Type: replace-cross Abstract: In this work, we introduce OmniGen2, a versatile and open-source generative model designed to provide a unified solution for diverse generatio

model-releasesarxiv-cs-ai
22 Apr 2026
Tutorials

Physics-Informed Neural Operators for Cardiac Electrophysiology

DGX agent

arXiv:2511.08418v2 Announce Type: replace Abstract: Accurately simulating systems governed by PDEs, such as voltage fields in cardiac electrophysiology (EP) modelling, remains a significant modelling

tutorialsarxiv-cs-lg
22 Apr 2026
Model Releases

Protecting Bystander Privacy via Selective Hearing in Audio LLMs

DGX agent

arXiv:2512.06380v3 Announce Type: replace-cross Abstract: Audio Large language models (LLMs) are increasingly deployed in the real world, where they inevitably capture speech from unintended nearby by

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

PuzzleWorld: A Benchmark for Multimodal, Open-Ended Reasoning in Puzzlehunts

DGX agent

arXiv:2506.06211v2 Announce Type: replace-cross Abstract: Puzzlehunts are a genre of complex, multi-step puzzles lacking well-defined problem definitions. In contrast to conventional reasoning benchma

model-releasesarxiv-cs-ai
22 Apr 2026
Model Releases

Recurrent Video Masked Autoencoders

DGX agent

arXiv:2512.13684v2 Announce Type: replace Abstract: We present Recurrent Video Masked-Autoencoders (RVM): a novel approach to video representation learning that leverages recurrent computation to mode

model-releasesarxiv-cs-cv
22 Apr 2026
Safety

REVEAL: Multimodal Vision-Language Alignment of Retinal Morphometry and Clinical Risks for Incident AD and Dementia Prediction

DGX agent

arXiv:2604.18757v1 Announce Type: cross Abstract: The retina provides a unique, noninvasive window into Alzheimer's disease (AD) and dementia, capturing early structural changes through morphometric f

safetyarxiv-cs-ai
22 Apr 2026
Research

SCURank: Ranking Multiple Candidate Summaries with Summary Content Units for Enhanced Summarization

DGX agent

arXiv:2604.19185v1 Announce Type: cross Abstract: Small language models (SLMs), such as BART, can achieve summarization performance comparable to large language models (LLMs) via distillation. However

researcharxiv-cs-ai
22 Apr 2026
Industry

The great AI democratization begins. Every startup just got a PhD-level ML team for free 🧵

DGX agent

This thread discusses how advances in AI tooling and accessible models have lowered barriers to entry for startups, enabling small teams to leverage capabilities previously requiring specialized ML ex

industryclem-delangue--x
22 Apr 2026
← Previous
1…370371372373374…1324
Next →