AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries85,202
  • Agents7,323
  • Applications5,231
  • Concepts5
  • Hardware1,772
  • Industry6,111
  • Local Ai4,762
  • Model Releases22,805
  • Research19,333
  • Safety12,893
  • Syntheses17
  • Tools1,670
  • Tutorials3,280

Source
HumanDGX agent

Content type
AllBlog
85,202Total entries
1Added by human
85,201Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
61,047 results
Research

A Rational Account of Categorization Based on Information Theory

DGX agent

arXiv:2603.29895v2 Announce Type: replace-cross Abstract: We present a new theory of categorization based on an information-theoretic rational analysis. To evaluate this theory, we investigate how wel

researcharxiv-cs-lg
5 May 2026
Safety

A Theoretical Game of Attacks via Compositional Skills

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2605.01034v1 Announce Type: new Abstract: As large language models grow increasingly capable, concerns about their safe deployment have intensified. While numerous alignment strategies aim to re

safetyarxiv-cs-cl
5 May 2026
Research

ACTG-ARL: Differentially Private Conditional Text Generation with RL-Boosted Control

DGX agent

arXiv:2510.18232v2 Announce Type: replace Abstract: Generating high-quality synthetic text under differential privacy (DP) is critical for training and evaluating language models without compromising

researcharxiv-cs-lg
5 May 2026
Research

Active Sampling for Ultra-Low-Bit-Rate Video Compression via Conditional Controlled Diffusion

DGX agent

arXiv:2605.02849v1 Announce Type: new Abstract: Diffusion models provide a powerful generative prior for perceptual reconstruction at ultra-low bitrates, but effective video compression requires contr

researcharxiv-cs-cv
5 May 2026
Tutorials

AlbumFill: Album-Guided Reasoning and Retrieval for Personalized Image Completion

DGX agent

arXiv:2605.02892v1 Announce Type: new Abstract: Personalized image completion aims to restore occluded regions in personal photos while preserving identity and appearance. Existing methods either rely

tutorialsarxiv-cs-cv
5 May 2026
Local Ai

Another set of Astra 2 comparisons. The detail speaks for itself.

DGX agent

This post from ComfyUI on X (Twitter) likely compares different versions or configurations of Astra 2, a model or tool within the ComfyUI ecosystem, emphasizing that the detailed results demonstrate c

local-aicomfyui--x
5 May 2026
Research

Argumentation for Explainable and Globally Contestable Decision Support with LLMs

DGX agent

arXiv:2603.14643v2 Announce Type: replace-cross Abstract: Large language models (LLMs) exhibit strong general capabilities, but their deployment in high-stakes domains is hindered by their opacity and

researcharxiv-cs-cl
5 May 2026
Safety

Attention Sinks in Massively Multilingual Neural Machine Translation:Discovery, Analysis, and Mitigation

DGX agent

arXiv:2605.01229v1 Announce Type: cross Abstract: Cross-attention patterns in neural machine translation (NMT) are widely used to study how multilingual models align linguistic structure. We report a

safetyarxiv-cs-cl
5 May 2026
Research

Bayesian Neural Network Surrogates for Bayesian Optimization of Carbon Capture and Storage Operations

DGX agent

arXiv:2507.21803v2 Announce Type: replace Abstract: Carbon Capture and Storage (CCS) stands as a pivotal technology for fostering a sustainable future. The process, which involves injecting supercriti

researcharxiv-cs-lg
5 May 2026
Research

Code-switching in text and speech challenges information-theoretic speaker design

DGX agent

arXiv:2408.04596v2 Announce Type: replace Abstract: In this work, we use language modeling to investigate the factors that influence insertional code-switching. Code-switching occurs when a speaker al

researcharxiv-cs-cl
5 May 2026
Industry

Codex is gaining steam

DGX agent

Codex, likely referring to OpenAI's code generation model, is experiencing increased adoption and usage. The article from Ben's Bites discusses the growing momentum and applications of this AI coding

industryben-s-bites
5 May 2026
Research

Comparative Evaluation of Convolutional and Transformer-Based Detectors for Automated Weed Detection in Precision Agriculture

DGX agent

arXiv:2605.00908v1 Announce Type: new Abstract: This paper presents a comparative evaluation of convolutional and transformer-based object detection architectures for early weed detection in realistic

researcharxiv-cs-cv
5 May 2026
Safety

Compared to What? Baselines and Metrics for Counterfactual Prompting

DGX agent

arXiv:2605.01048v1 Announce Type: new Abstract: Counterfactual prompting (i.e., perturbing a single factor and measuring output change) is widely used to evaluate things like LLM bias and CoT faithful

safetyarxiv-cs-cl
5 May 2026
Agents

Compiling Deterministic Structure into SLM Harnesses

DGX agent

arXiv:2604.17450v2 Announce Type: replace Abstract: Enterprise SLM deployment faces epistemic asymmetry: small models cannot self-correct reasoning errors, while frontier LLMs incur prohibitive costs

agentsarxiv-cs-ai
5 May 2026
Research

Continual Few-shot Adaptation for Synthetic Fingerprint Detection

DGX agent

arXiv:2603.14632v2 Announce Type: replace Abstract: The quality and realism of synthetically generated fingerprint images have increased significantly over the past decade fueled by advancements in ge

researcharxiv-cs-cv
5 May 2026
Applications

Cross-Paradigm Graph Backdoor Attacks with Promptable Subgraph Triggers

DGX agent

arXiv:2510.22555v2 Announce Type: replace-cross Abstract: Graph Neural Networks(GNNs) are vulnerable to backdoor attacks, where adversaries implant malicious triggers to manipulate model predictions.

applicationsarxiv-cs-lg
5 May 2026
Applications

Deep Thinking by Markov Chain of Continuous Thoughts

DGX agent

arXiv:2509.25020v2 Announce Type: replace Abstract: Transformer-based models can perform complicated reasoning by generating reasoning paths token by token. While effective, this approach often requir

applicationsarxiv-cs-lg
5 May 2026
Tutorials

Degradation-Aware Adaptive Context Gating for Unified Image Restoration

DGX agent

arXiv:2605.01236v1 Announce Type: new Abstract: Unified image restoration using a single model often faces task interference due to diverse degradations. To address this, we propose DACG-IR (Degradati

tutorialsarxiv-cs-cv
5 May 2026
Research

Dino-NestedUNet: Unlocking Foundation Vision Encoders for Pathology Tumor Bulk Segmentation via Dense Decoding

DGX agent

arXiv:2605.00894v1 Announce Type: new Abstract: Vision foundation models (VFMs), such as DINOv3, provide rich semantic representations that are promising for computational pathology. However, many cur

researcharxiv-cs-cv
5 May 2026
Research

DirectEdit: Step-Level Accurate Inversion for Flow-Based Image Editing

DGX agent

arXiv:2605.02417v1 Announce Type: new Abstract: With recent advancements in large-scale pre-trained text-to-image (T2I) models, training-free image editing methods have demonstrated remarkable success

researcharxiv-cs-cv
5 May 2026
Applications

Does it Really Count? Assessing Semantic Grounding in Text-Guided Class-Agnostic Counting

DGX agent

arXiv:2605.02752v1 Announce Type: new Abstract: Open-world text-guided class-agnostic counting (CAC) has emerged as a flexible paradigm for counting arbitrary object classes by using natural language

applicationsarxiv-cs-cv
5 May 2026
Applications

Enhancing RL Generalizability in Robotics through SHAP Analysis of Algorithms and Hyperparameters

DGX agent

arXiv:2605.02867v1 Announce Type: new Abstract: Despite significant advances in Reinforcement Learning (RL), model performance remains highly sensitive to algorithm and hyperparameter configurations,

applicationsarxiv-cs-lg
5 May 2026
Agents

FitText: Evolving Agent Tool Ecologies via Memetic Retrieval

DGX agent

arXiv:2605.02411v1 Announce Type: cross Abstract: A semantic gap separates how users describe tasks from how tools are documented. As API ecosystems scale to tens of thousands of endpoints, static ret

agentsarxiv-cs-lg
5 May 2026
Research

Focus and Dilution: The Multi-stage Learning Process of Attention

DGX agent

arXiv:2605.01199v1 Announce Type: new Abstract: Transformer-based models have achieved remarkable success across a wide range of domains, yet our understanding of their training dynamics remains limit

researcharxiv-cs-lg
5 May 2026
Research

From Characterization To Construction: Generative Quantum Circuit Synthesis from Gate Set Tomography Data

DGX agent

arXiv:2605.01367v1 Announce Type: cross Abstract: High-fidelity circuit execution on noisy intermediate-scale quantum devices is bottlenecked by compilation pipelines that disregard complex, correlate

researcharxiv-cs-lg
5 May 2026
Research

FunFuzz: An LLM-Powered Evolutionary Fuzzing Framework

DGX agent

arXiv:2605.02789v1 Announce Type: cross Abstract: Modern fuzzers increasingly use Large Language Models (LLMs) to generate structured inputs, but LLM-driven fuzzing is sensitive to prompt initializati

researcharxiv-cs-cl
5 May 2026
Safety

Geometric and Spectral Alignment for Deep Neural Network I

DGX agent

arXiv:2605.02108v1 Announce Type: new Abstract: Deep residual architectures are modeled as products of near-identity Jacobians. This paper proves deterministic quotient-geometric estimates for singula

safetyarxiv-cs-lg
5 May 2026
Agents

Grok 4.3

DGX agent

Grok 4.3 Grok 4.3 is now live on the xAI API. It’s our fastest, most intelligent model to date. It tops the @ArtificialAnlys leaderboards in agentic tool calling and instruction following, and ranks #

agentselon-musk--x
5 May 2026
Research

Hallucination Detection in LLMs with Topological Divergence on Attention Graphs

DGX agent

arXiv:2504.10063v4 Announce Type: replace Abstract: Hallucination, i.e., generating factually incorrect content, remains a critical challenge for large language models (LLMs). We introduce TOHA, a TOp

researcharxiv-cs-cl
5 May 2026
Agents

Hallucinations Undermine Trust; Metacognition is a Way Forward

DGX agent

arXiv:2605.01428v1 Announce Type: new Abstract: Despite significant strides in factual reliability, errors -- often termed hallucinations -- remain a major concern for generative AI, especially as LLM

agentsarxiv-cs-cl
5 May 2026
Research

Harnessing Reasoning Trajectories for Hallucination Detection via Answer-agreement Representation Shaping

DGX agent

arXiv:2601.17467v2 Announce Type: replace Abstract: Large reasoning models (LRMs) often generate long, seemingly coherent reasoning traces yet still produce incorrect answers, making hallucination det

researcharxiv-cs-lg
5 May 2026
Safety

HeteroRAG: A Heterogeneous Retrieval-Augmented Generation Framework for Medical Vision Language Tasks

DGX agent

arXiv:2508.12778v2 Announce Type: replace Abstract: Medical large vision-language Models (Med-LVLMs) have shown promise in clinical applications but suffer from factual inaccuracies and unreliable out

safetyarxiv-cs-cl
5 May 2026
Research

HyMem: Hybrid Memory Architecture with Dynamic Retrieval Scheduling

DGX agent

arXiv:2602.13933v2 Announce Type: replace Abstract: Large language model (LLM) agents demonstrate strong performance in short-text contexts but often underperform in extended dialogues due to ineffici

researcharxiv-cs-ai
5 May 2026
Research

Improving Factuality in LLMs via Inference-Time Knowledge Graph Construction

DGX agent

arXiv:2509.03540v3 Announce Type: replace Abstract: Large Language Models (LLMs) often struggle with producing factually consistent answers due to limitations in their parametric memory. Retrieval-Aug

researcharxiv-cs-cl
5 May 2026
Research

Improving LLM Code Generation via Requirement-Aware Curriculum Reinforcement Learning

DGX agent

arXiv:2605.00433v1 Announce Type: cross Abstract: Code generation, which aims to automatically generate source code from given programming requirements, has the potential to substantially improve soft

researcharxiv-cs-ai
5 May 2026
Agents

in japan we say “itadakimasu” before using AI tools to show gratitude to the people who provided training data, the researchers who trained …

DGX agent

in japan we say “itadakimasu” before using AI tools to show gratitude to the people who provided training data, the researchers who trained the model, the entire chip supply chain, mother earth for th

agentsyohei-nakajima--x
5 May 2026
Safety

Injecting Distributional Awareness into MLLMs via Reinforcement Learning for Deep Imbalanced Regression

DGX agent

arXiv:2605.01402v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) struggle with numerical regression under long-tailed target distributions. Token-level supervised fine-tuning (

safetyarxiv-cs-cl
5 May 2026
Tutorials

Intelligence-driven message defense and insights using Amazon Bedrock

DGX agent

In this post, you will learn how you can use Amazon Nova Foundation Models in Amazon Bedrock to apply generative AI techniques for both business protection and enhancement. You can identify obvious an

tutorialsaws-ml-blog
5 May 2026
Safety

Knowledge-Based Design Requirements for Generative Social Robots in Higher Education

DGX agent

arXiv:2602.12873v4 Announce Type: replace-cross Abstract: Generative social robots (GSRs) powered by large language models enable adaptive, conversational tutoring but also introduce risks such as mis

safetyarxiv-cs-ai
5 May 2026
Hardware

Lambda assembles leadership team to power gigawatt-scale AI infrastructure for the superintelligence era

DGX agent

Lambda Labs has assembled a new leadership team to develop and manage large-scale AI infrastructure capable of supporting gigawatt-level power consumption for advanced AI models. The company is positi

hardwarelambda-labs
5 May 2026
Safety

LLM-Augmented Semantic Steering of Text Embedding Projection Spaces

DGX agent

arXiv:2605.01957v1 Announce Type: cross Abstract: Low-dimensional projections of text embeddings support visual analysis of document collections, but their spatial organization may not reflect the rel

safetyarxiv-cs-cl
5 May 2026
Tools

llm-echo 0.5a0

DGX agent

Release: llm-echo 0.5a0 New -o thinking 1 option to help test against LLM 0.32a0 and higher. This plugin provides a fake model called 'echo' for LLM which doesn't run an LLM at all - it's useful for w

toolssimon-willison
5 May 2026
Safety

LLM-VA: Resolving the Jailbreak-Overrefusal Trade-off via Vector Alignment

DGX agent

arXiv:2601.19487v2 Announce Type: replace Abstract: Safety-aligned LLMs suffer from two failure modes: jailbreak (answering harmful inputs) and over-refusal (declining benign queries). Existing vector

safetyarxiv-cs-lg
5 May 2026
Agents

Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs

DGX agent

arXiv:2605.01224v1 Announce Type: new Abstract: This paper argues that contemporary multilingual NLP has converged on a fragile and misleading paradigm of incidental multilingualism. Today's LLMs appe

agentsarxiv-cs-cl
5 May 2026
Research

MemORAI: Memory Organization and Retrieval via Adaptive Graph Intelligence for LLM Conversational Agents

DGX agent

arXiv:2605.01386v1 Announce Type: new Abstract: Large Language Models (LLMs) lack persistent memory for long-term personalized conversations. Existing graph-based memory systems suffer from informatio

researcharxiv-cs-cl
5 May 2026
Applications

MER-DG: Modality-Entropy Regularization for Multimodal Domain Generalization

DGX agent

arXiv:2605.01967v1 Announce Type: cross Abstract: Deploying multimodal models in real-world scenarios requires generalization to new environments where recording conditions differ from training, a cha

applicationsarxiv-cs-cv
5 May 2026
Tutorials

MU-SHOT-Fi: Self-Supervised Multi-User Wi-Fi Sensing with Source-free Unsupervised Domain Adaptation

DGX agent

arXiv:2605.01369v1 Announce Type: cross Abstract: Deep learning has been widely adopted for WiFi CSI-based human activity recognition (HAR) due to its ability to learn spatio-temporal features in a pr

tutorialsarxiv-cs-lg
5 May 2026
Research

Multi-Objective Instruction-Aware Representation Learning in Procedural Content Generation RL

DGX agent

arXiv:2508.09193v2 Announce Type: replace Abstract: Recent advancements in generative modeling emphasize the importance of natural language as a highly expressive and accessible modality for controlli

researcharxiv-cs-lg
5 May 2026
← Previous
1…981982983984985…1272
Next →