AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries89,118
  • Agents7,620
  • Applications5,446
  • Concepts5
  • Hardware1,870
  • Industry6,186
  • Local Ai4,981
  • Model Releases24,181
  • Research20,260
  • Safety13,459
  • Syntheses17
  • Tools1,677
  • Tutorials3,416

Source
HumanDGX agent

89,118Total entries
1Added by human
89,117Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
64,222 results
16 Apr 2026

Built an political benchmark for LLMs. KIMI K2 can't answer about Taiwan (Obviously). GPT-5.3 refuses 100% of questions when given an opt-out. [P]

Model ReleasesDGX agent

A researcher on r/MachineLearning built a political benchmark to evaluate how various LLMs handle sensitive geopolitical and politically contentious questions. Key findings include that Kimi K2 (Moons

ChartNet: A Million-Scale, High-Quality Multimodal Dataset for Robust Chart Understanding

SafetyDGX agent

arXiv:2603.27064v2 Announce Type: replace-cross Abstract: Understanding charts requires models to jointly reason over geometric visual patterns, structured numerical data, and natural language -- a ca

Coding agents learn from experience, but that knowledge stays locked in silos. Solve a thousand SWE tasks, and none of that wisdom helps wit…

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tutorials
DGX agent

Coding agents learn from experience, but that knowledge stays locked in silos. Solve a thousand SWE tasks, and none of that wisdom helps with competitive coding. What if memories could transfer across

Deep Learning Based Amharic Chatbot for FAQs in Universities

ResearchDGX agent

arXiv:2402.01720v3 Announce Type: replace-cross Abstract: University students often spend a considerable amount of time seeking answers to common questions from administrators or teachers. This can be

Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage

Model ReleasesDGX agent

arXiv:2604.13060v1 Announce Type: new Abstract: Dental triage is a safety-critical clinical routing task that requires integrating multimodal clinical information (e.g., patient complaints and radiogr

Design Conditions for Intra-Group Learning of Sequence-Level Rewards: Token Gradient Cancellation

ResearchDGX agent

arXiv:2604.13088v1 Announce Type: new Abstract: In sparse termination rewards, intra-group comparisons have become the dominant paradigm for fine-tuning reasoning models via reinforcement learning. Ho

Design Space Exploration of Hybrid Quantum Neural Networks for Chronic Kidney Disease

Model ReleasesDGX agent

arXiv:2604.13608v1 Announce Type: new Abstract: Hybrid Quantum Neural Networks (HQNNs) have recently emerged as a promising paradigm for near-term quantum machine learning. However, their practical pe

EMGFlow: Robust and Efficient Surface Electromyography Synthesis via Flow Matching

Model ReleasesDGX agent

arXiv:2604.13685v1 Announce Type: cross Abstract: Deep learning-based surface electromyography (sEMG) gesture recognition is frequently bottlenecked by data scarcity and limited subject diversity. Whi

Evaluating the Evaluator: Problems with SemEval-2020 Task 1 for Lexical Semantic Change Detection

Model ReleasesDGX agent

arXiv:2604.13232v1 Announce Type: new Abstract: This discussion paper re-examines SemEval-2020 Task 1, the most influential shared benchmark for lexical semantic change detection, through a three-part

From Where Words Come: Efficient Regularization of Code Tokenizers Through Source Attribution

SafetyDGX agent

arXiv:2604.14053v1 Announce Type: new Abstract: Efficiency and safety of Large Language Models (LLMs), among other factors, rely on the quality of tokenization. A good tokenizer not only improves infe

Hierarchical Reinforcement Learning with Augmented Step-Level Transitions for LLM Agents

ResearchDGX agent

arXiv:2604.05808v2 Announce Type: replace-cross Abstract: Large language model (LLM) agents have demonstrated strong capabilities in complex interactive decision-making tasks. However, existing LLM ag

Hybrid Approach for Enhancing Lesion Segmentation in Fundus Images

ResearchDGX agent

arXiv:2509.25549v2 Announce Type: replace Abstract: Choroidal nevi are common benign pigmented lesions in the eye, with a small risk of transforming into melanoma. Early detection is critical to impro

I edited the intro because I realized I buried the lede originally- The 1M context window is a double-edged sword. It allows Claude to do mo…

Model ReleasesDGX agent

I edited the intro because I realized I buried the lede originally- The 1M context window is a double-edged sword. It allows Claude to do more complex tasks but it can also leads to more context pollu

I tested Ernie Image Turbo (fp8, nvfp4, fp16 and INT8) with Nano Banana Pro 2 Prompts so you won't have to

Local AiDGX agent

This r/StableDiffusion post is a community-driven benchmark comparing multiple quantization formats — fp8, nvfp4, fp16, and INT8 — of Baidu's ERNIE Image Turbo text-to-image model, using a standardize

Joint Representation Learning and Clustering via Gradient-Based Manifold Optimization

Model ReleasesDGX agent

arXiv:2604.13484v1 Announce Type: cross Abstract: Clustering and dimensionality reduction have been crucial topics in machine learning and computer vision. Clustering high-dimensional data has been ch

Learning Sewing Patterns via Latent Flow Matching of Implicit Fields

ResearchDGX agent

arXiv:2601.17740v2 Announce Type: replace Abstract: Sewing patterns define the structural foundation of garments and are essential for applications such as fashion design, fabrication, and physical si

Leveraging LLM-GNN Integration for Open-World Question Answering over Knowledge Graphs

Model ReleasesDGX agent

arXiv:2604.13979v1 Announce Type: new Abstract: Open-world Question Answering (OW-QA) over knowledge graphs (KGs) aims to answer questions over incomplete or evolving KGs. Traditional KGQA assumes a c

LiteParse hit 4.3K+ GitHub stars in a few weeks. Today it officially joins the LlamaIndex ecosystem, with its own page at http://www.llamain…

Model ReleasesDGX agent

LiteParse hit 4.3K+ GitHub stars in a few weeks. Today it officially joins the LlamaIndex ecosystem, with its own page at http://www.llamaindex.ai/liteparse?utm_medium=socials&utm_source=twitter&utm_c

LiteParse should be the default document parser you use with any AI agent (Claude Code, Claude Cowork, OpenClaw, Codex, and more) The core i…

Model ReleasesDGX agent

LiteParse should be the default document parser you use with any AI agent (Claude Code, Claude Cowork, OpenClaw, Codex, and more) The core is extremely fast text and accurate parsing from any document

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding

Model ReleasesDGX agent

arXiv:2602.20913v2 Announce Type: replace Abstract: This paper addresses the critical and underexplored challenge of long video understanding with low computational budgets. We propose LongVideo-R1, a

Olfactory pursuit: catching a moving odor source in complex flows

Model ReleasesDGX agent

arXiv:2604.13121v1 Announce Type: new Abstract: Locating and intercepting a moving target from possibly delayed, intermittent sensory signals is a paradigmatic problem in decision-making under uncerta

Parameter-efficient Quantum Multi-task Learning

Model ReleasesDGX agent

arXiv:2604.13560v1 Announce Type: new Abstract: Multi-task learning (MTL) improves generalization and data efficiency by jointly learning related tasks through shared representations. In the widely us

PAT-VCM: Plug-and-Play Auxiliary Tokens for Video Coding for Machines

ResearchDGX agent

arXiv:2604.13294v1 Announce Type: new Abstract: Existing video coding for machines is often trained for a specific downstream task and model. As a result, the compressed representation becomes tightly

Physics-Informed Neural Networks for Methane Sorption: Cross-Gas Transfer Learning, Ensemble Collapse Under Physics Constraints, and Monte Carlo Dropout Uncertainty Quantification

ResearchDGX agent

arXiv:2604.13992v1 Announce Type: new Abstract: Accurate methane sorption prediction across heterogeneous coal ranks requires models that combine thermodynamic consistency, efficient knowledge transfe

Predicting Time Pressure of Powered Two-Wheeler Riders for Proactive Safety Interventions

Model ReleasesDGX agent

arXiv:2601.03173v2 Announce Type: replace Abstract: Time pressure critically influences risky maneuvers and crash proneness among powered two-wheeler riders, yet its prediction remains underexplored i

qwen3.6 is out

Local AiDGX agent

Qwen 3.6 Plus Preview is Alibaba's next-generation large language model released on March 30-31, 2026 , and the first open-weight variant was released following the February Qwen 3.5 series, prioritiz

UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry

Model ReleasesDGX agent

arXiv:2604.13584v1 Announce Type: new Abstract: We present UNRIO, an uncertainty-aware radar-inertial odometry system that estimates ego-velocity directly from raw mmWave radar IQ signals rather than

Unsupervised Anomaly Detection in Process-Complex Industrial Time Series: A Real-World Case Study

Model ReleasesDGX agent

arXiv:2604.13928v1 Announce Type: new Abstract: Industrial time-series data from real production environments exhibits substantially higher complexity than commonly used benchmark datasets, primarily

Vision-and-Language Navigation for UAVs: Progress, Challenges, and a Research Roadmap

AgentsDGX agent

arXiv:2604.13654v1 Announce Type: new Abstract: Vision-and-Language Navigation for Unmanned Aerial Vehicles (UAV-VLN) represents a pivotal challenge in embodied artificial intelligence, focused on ena

Visual Self-Fulfilling Alignment: Shaping Safety-Oriented Personas via Threat-Related Images

SafetyDGX agent

arXiv:2603.08486v2 Announce Type: replace Abstract: Multimodal large language models (MLLMs) face safety misalignment, where visual inputs enable harmful outputs. To address this, existing methods req

Visual Sparse Steering (VS2): Unsupervised Adaptation for Image Classification using Sparsity-Guided Steering Vectors

ResearchDGX agent

arXiv:2506.01247v2 Announce Type: replace Abstract: Steering vision foundation models at test time, without updating foundation-model weights or using labeled target data, is a desirable yet challengi

What Are We Really Measuring? Rethinking Dataset Bias in Web-Scale Natural Image Collections via Unsupervised Semantic Clustering

SafetyDGX agent

arXiv:2604.13610v1 Announce Type: new Abstract: In computer vision, a prevailing method for quantifying dataset bias is to train a model to distinguish between datasets. High classification accuracy i

What’s new with Google Data Cloud

Model ReleasesDGX agent

April 13 - April 17 We announced we are reintroducing Data Studio to play a significant role in the AI era, expanding from data visualizations and reports to host BigQuery conversational agents and da

Why MLLMs Struggle to Determine Object Orientations

SafetyDGX agent

arXiv:2604.13321v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) struggle with tasks that require reasoning about 2D object orientation in images, as documented in prior work.

15 Apr 2026

A Benchmark for Evaluating Outcome-Driven Constraint Violations in Autonomous AI Agents

Model ReleasesDGX agent

arXiv:2512.20798v4 Announce Type: replace Abstract: As autonomous AI agents are deployed in high-stakes environments, ensuring their safety has become a paramount concern. Existing safety benchmarks p

A Large-Scale Comparative Analysis of Imputation Methods for Single-Cell RNA Sequencing Data

Model ReleasesDGX agent

arXiv:2603.24626v2 Announce Type: replace-cross Abstract: Background: Single-cell RNA sequencing (scRNA-seq) enables gene expression profiling at cellular resolution but is inherently affected by spar

AAPO: Enhancing the Reasoning Capabilities of LLMs with Advantage Margin

SafetyDGX agent

arXiv:2505.14264v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has emerged as an effective approach for enhancing the reasoning capabilities of large language models (LLMs), esp

Active Imitation Learning for Thermal- and Kernel-Aware LFM Inference on 3D S-NUCA Many-Cores

SafetyDGX agent

arXiv:2604.11948v1 Announce Type: new Abstract: Large Foundation Model (LFM) inference is both memory- and compute-intensive, traditionally relying on GPUs. However, the limited availability and high

Adobe takes Creative Cloud into Claude Code-esque territory

Model ReleasesDGX agent

Adobe has introduced the **Firefly AI Assistant**, a new agentic tool that brings autonomous, multi-step workflow capabilities to Creative Cloud — drawing comparisons to AI coding agents like Claude C

Aethon: A Reference-Based Replication Primitive for Constant-Time Instantiation of Stateful AI Agents

AgentsDGX agent

arXiv:2604.12129v1 Announce Type: new Abstract: The transition from stateless model inference to stateful agentic execution is reshaping the systems assumptions underlying modern AI infrastructure. Wh

AGSC: Adaptive Granularity and Semantic Clustering for Uncertainty Quantification in Long-text Generation

ResearchDGX agent

arXiv:2604.06812v2 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated impressive capabilities in long-form generation, yet their application is hindered by the hallucinati

[AINews] Humanity's Last Gasp

ToolsDGX agent

'Humanity's Last Gasp' is an AI news roundup from the Latent Space newsletter, likely covering significant developments in AI safety, existential risk discussions, or major industry milestones that pr

AlphaEval: Evaluating Agents in Production

Model ReleasesDGX agent

arXiv:2604.12162v1 Announce Type: new Abstract: The rapid deployment of AI agents in commercial settings has outpaced the development of evaluation methodologies that reflect production realities. Exi

Anima Dataset…issues.

Local AiDGX agent

This Reddit thread on r/StableDiffusion likely discusses user-reported problems and troubleshooting related to the training dataset for the Anima model — a 2-billion-parameter text-to-image model by C

Any good settings for Character lora for Klein?

Local AiDGX agent

This Reddit post from r/StableDiffusion seeks community advice on optimal generation settings for using a character LoRA built on the FLUX.2 Klein model, which is a compact 4B–9B parameter model suite

ASTRA: Let Arbitrary Subjects Transform in Video Editing

Model ReleasesDGX agent

arXiv:2510.01186v2 Announce Type: replace Abstract: While existing video editing methods excel with single subjects, they struggle in dense, multi-subject scenes, frequently suffering from attention d

Beyond Majority Voting: Efficient Best-Of-N with Radial Consensus Score

AgentsDGX agent

arXiv:2604.12196v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate multiple candidate responses for a given prompt, yet selecting the most reliable one remains challengin

BID-LoRA: A Parameter-Efficient Framework for Continual Learning and Unlearning

Model ReleasesDGX agent

arXiv:2604.12686v1 Announce Type: cross Abstract: Recent advances in deep learning underscore the need for systems that can not only acquire new knowledge through Continual Learning (CL) but also remo

Built an open-source local-AI resume tailoring app called RoleCraft

Local AiDGX agent

RoleCraft is an open-source, privacy-focused resume tailoring application that runs AI models locally using Ollama, allowing users to customize their resumes to match specific job descriptions without

Cline with Ollama on a RTX4090 (24GRAM) and i9 with 64 GRAM

Local AiDGX agent

This Reddit post from r/ollama discusses a user's experience running Cline (an AI coding agent) with Ollama on a high-end local hardware setup consisting of an NVIDIA RTX 4090 with 24GB VRAM and an In

CODESTRUCT: Code Agents over Structured Action Spaces

Model ReleasesDGX agent

arXiv:2604.05407v2 Announce Type: replace Abstract: LLM-based code agents treat repositories as unstructured text, applying edits through brittle string matching that frequently fails due to formattin

Combating Pattern and Content Bias: Adversarial Feature Learning for Generalized AI-Generated Image Detection

SafetyDGX agent

arXiv:2604.12353v1 Announce Type: new Abstract: In recent years, the rapid development of generative artificial intelligence technology has significantly lowered the barrier to creating high-quality f

Compiling Activation Steering into Weights via Null-Space Constraints for Stealthy Backdoors

SafetyDGX agent

arXiv:2604.12359v1 Announce Type: cross Abstract: Safety-aligned large language models (LLMs) are increasingly deployed in real-world pipelines, yet this deployment also enlarges the supply-chain atta

CropVLM: Learning to Zoom for Fine-Grained Vision-Language Perception

ResearchDGX agent

arXiv:2511.19820v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) often struggle with tasks that require fine-grained image understanding, such as scene-text recognition or docum

Cross-Attentive Multiview Fusion of Vision-Language Embeddings

ResearchDGX agent

arXiv:2604.12551v1 Announce Type: new Abstract: Vision-language models have been key to the development of open-vocabulary 2D semantic segmentation. Lifting these models from 2D images to 3D scenes, h

Cross-Cultural Simulation of Citizen Emotional Responses to Bureaucratic Red Tape Using LLM Agents

SafetyDGX agent

arXiv:2604.12545v1 Announce Type: new Abstract: Improving policymaking is a central concern in public administration. Prior human subject studies reveal substantial cross-cultural differences in citiz

DocSeeker: Structured Visual Reasoning with Evidence Grounding for Long Document Understanding

SafetyDGX agent

arXiv:2604.12812v1 Announce Type: new Abstract: Existing Multimodal Large Language Models (MLLMs) suffer from significant performance degradation on the long document understanding task as document le

Every Picture Tells a Dangerous Story: Memory-Augmented Multi-Agent Jailbreak Attacks on VLMs

SafetyDGX agent

arXiv:2604.12616v1 Announce Type: new Abstract: The rapid evolution of Vision-Language Models (VLMs) has catalyzed unprecedented capabilities in artificial intelligence; however, this continuous modal

FeaXDrive: Feasibility-aware Trajectory-Centric Diffusion Planning for End-to-End Autonomous Driving

Model ReleasesDGX agent

arXiv:2604.12656v1 Announce Type: cross Abstract: End-to-end diffusion planning has shown strong potential for autonomous driving, but the physical feasibility of generated trajectories remains insuff

Fragile Reconstruction: Adversarial Vulnerability of Reconstruction-Based Detectors for Diffusion-Generated Images

SafetyDGX agent

arXiv:2604.12781v1 Announce Type: new Abstract: Recently, detecting AI-generated images produced by diffusion-based models has attracted increasing attention due to their potential threat to safety. A

← Previous
1…495496497498499…1071
Next →