AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,648
  • Agents7,273
  • Applications5,201
  • Concepts5
  • Hardware1,758
  • Industry6,104
  • Local Ai4,732
  • Model Releases22,612
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,669
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,648Total entries
1Added by human
84,647Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,587 results
Research

From Local to Global to Mechanistic: An iERF-Centered Unified Framework for Interpreting Vision Models

DGX agent

arXiv:2605.00474v1 Announce Type: new Abstract: Modern vision models achieve remarkable accuracy, but explaining where evidence arises, what the model encodes, and how internal computations assemble t

researcharxiv-cs-cv
4 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Probing Multimodal Large Language Models on Cognitive Biases in Chinese Short-Video Misinformation

DGX agent

arXiv:2601.06600v2 Announce Type: replace Abstract: Short-video platforms have become major channels for misinformation, where deceptive claims frequently leverage visual experiments and social cues.

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

SC-Taxo: Hierarchical Taxonomy Generation under Semantic Consistency Constraints using Large Language Models

DGX agent

arXiv:2605.00620v1 Announce Type: new Abstract: Scientific literature is expanding at an unprecedented pace, making it increasingly challenging to efficiently organize and access domain knowledge. A h

model-releasesarxiv-cs-cl
4 May 2026
Research

Trees to Flows and Back: Unifying Decision Trees and Diffusion Models

DGX agent

arXiv:2605.00414v1 Announce Type: new Abstract: Decision trees and diffusion models are ostensibly disparate model classes, one discrete and hierarchical, the other continuous and dynamic. This work u

researcharxiv-cs-lg
4 May 2026
Agents

Trinity-Large-Thinking, @arcee_ai's latest model, is now free on Nous Portal for the next week Sign up for Nous Portal to use it in your Her…

DGX agent

Nous Research announced that Trinity-Large-Thinking, a new model from Arcee AI, is available for free on the Nous Portal for a limited one-week period. Users can access the model by signing up for the

agentsnous-research--x
4 May 2026
Applications

This is a good explanation of why the gap between open and closed models is larger than it appears in benchmarks. I would add in that curren…

DGX agent

This is a good explanation of why the gap between open and closed models is larger than it appears in benchmarks. I would add in that current open models are also more fragile than closed: they handle

applicationsethan-mollick--x
3 May 2026
Safety

Analytical Correction for Subsampling Bias in Drifting Models

DGX agent

arXiv:2604.27239v1 Announce Type: new Abstract: Drifting models are capable one-step generative models trained to follow a drifting field. The field combines attractive and repulsive softmax-weighted

safetyarxiv-cs-lg
1 May 2026
Safety

CoAX: Cognitive-Oriented Attribution eXplanation User Model of Human Understanding of AI Explanations

DGX agent

arXiv:2604.27354v1 Announce Type: new Abstract: Explainable AI (XAI) aims to improve user understanding and decisions when using AI models. However, despite innovations in XAI, recent user evaluations

safetyarxiv-cs-ai
1 May 2026
Safety

Debiasing Reward Models via Causally Motivated Inference-Time Intervention

DGX agent

arXiv:2604.27495v1 Announce Type: cross Abstract: Reward models (RMs) play a central role in aligning large language models (LLMs) with human preferences. However, RMs are often sensitive to spurious

safetyarxiv-cs-ai
1 May 2026
Research

Geometry-Calibrated Conformal Abstention for Language Models

DGX agent

arXiv:2604.27914v1 Announce Type: new Abstract: When language models lack relevant knowledge for a given query, they frequently generate plausible responses that can be hallucinations, rather than adm

researcharxiv-cs-cl
1 May 2026
Safety

“Marcus’ specific point about coding is structurally important: a model that produces code which compiles and passes the tests it was given …

DGX agent

“Marcus’ specific point about coding is structurally important: a model that produces code which compiles and passes the tests it was given is not the same as a model that produces correct, secure, ma

safetygary-marcus--x
1 May 2026
Safety

Policy-Grounded Safety Evaluation of 20 Large Language Models

DGX agent

arXiv:2507.14719v2 Announce Type: replace Abstract: As large language models (LLMs) become increasingly integrated into real-world applications, scalable and rigorous safety evaluation is essential. T

safetyarxiv-cs-ai
1 May 2026
Research

Rethinking Pulmonary Embolism Segmentation: A Study of Current Approaches and Challenges with an Open Weight Model

DGX agent

arXiv:2509.18308v3 Announce Type: replace Abstract: Pulmonary Embolism (PE) is a life-threatening condition for which accurate and timely detection is critical to patient care. However, our systematic

researcharxiv-cs-cv
1 May 2026
Local Ai

Seasoned dev but new to local LLMs: help me pick the right Apple product for hosting model in the 27B - 36B size

DGX agent

A seasoned developer seeking advice on selecting an Apple product to locally host large language models in the 27-36 billion parameter range, discussing the trade-offs and specifications of different

local-air-ollama
1 May 2026
Model Releases

Creating highly efficient agents: 450M tool-calling tokens distilled for post-training from top open-source models

DGX agent

Harnesses If you've used Claude Code or Codex, you've used a harness. A harness is the infrastructure layer that wraps an AI coding agent and decides how it operates, what it can touch, and how you me

model-releaseslambda-labs
30 Apr 2026
Safety

Data-Centric Foundation Models in Computational Healthcare: A Survey

DGX agent

arXiv:2401.02458v3 Announce Type: replace-cross Abstract: The advent of foundation models (FMs) as an emerging suite of AI techniques has struck a wave of opportunities in computational healthcare. Th

safetyarxiv-cs-ai
30 Apr 2026
Safety

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how y…

DGX agent

For better or worse, regulation for closed-source models served by a few (quite large) companies is easy. It is not as easy to imagine how you regulate open-source models that can be served by a range

safetyethan-mollick--x
30 Apr 2026
Model Releases

Information Extraction from Electricity Invoices with General-Purpose Large Language Models

DGX agent

arXiv:2604.25927v1 Announce Type: new Abstract: Information extraction from semi-structured business documents remains a critical challenge for enterprise management. This study evaluates the capabili

model-releasesarxiv-cs-cl
30 Apr 2026
Tools

Introducing Silico: the platform for building AI models with the precision of written software. Silico lets researchers and engineers see in…

DGX agent

Introducing Silico: the platform for building AI models with the precision of written software. Silico lets researchers and engineers see inside their models, debug failures, and intentionally design

toolslinus-lee--x
30 Apr 2026
Tutorials

Language Diffusion Models are Associative Memories Capable of Retrieving Unseen Data

DGX agent

arXiv:2604.26841v1 Announce Type: cross Abstract: When do language diffusion models memorize their training data, and how to quantitatively assess their true generative regime? We address these questi

tutorialsarxiv-cs-ai
30 Apr 2026
Safety

Sociodemographic Biases in Educational Counselling by Large Language Models

DGX agent

arXiv:2604.25932v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into educational settings, understanding their potential biases is critical. This study ex

safetyarxiv-cs-ai
30 Apr 2026
Safety

STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation

DGX agent

arXiv:2604.26848v1 Announce Type: new Abstract: Robotic manipulation critically requires reasoning about future spatial-temporal interactions, yet existing VLA policies and world-model-enhanced polici

safetyarxiv-cs-ro
30 Apr 2026
Industry

The Tesla Model X was the fastest-selling used vehicle in the U.S. in March 2026, according to a new study by iSeeCars. On average, it found…

DGX agent

The Tesla Model X was the fastest-selling used vehicle in the U.S. in March 2026, according to a new study by iSeeCars. On average, it found a buyer in 26 days, compared to the typical car, which usua

industryelon-musk--x
30 Apr 2026
Applications

A multi-stage soft computing framework for complex disease modelling and decision support: A liver cirrhosis case study

DGX agent

arXiv:2604.24796v1 Announce Type: cross Abstract: Liver cirrhosis is a major global health problem causing millions of deaths annually, and timely detection with aggressive treatment can significantly

applicationsarxiv-cs-lg
29 Apr 2026
Model Releases

ADE: Adaptive Dictionary Embeddings -- Scaling Multi-Anchor Representations to Large Language Models

DGX agent

arXiv:2604.24940v1 Announce Type: new Abstract: Word embeddings are fundamental to natural language processing, yet traditional approaches represent each word with a single vector, creating representa

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

After switching into the new directory (`cd lc-docs`) I'm going to set the name of my agent, and the model to use This is done in deepagents…

DGX agent

After switching into the new directory (`cd lc-docs`) I'm going to set the name of my agent, and the model to use This is done in deepagents.toml I'm going to use @Zai_org GLM5 model, served via @base

agentsharrison-chase--x
29 Apr 2026
Model Releases

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

DGX agent

arXiv:2602.15547v2 Announce Type: replace Abstract: Text embedding models are widely used for semantic similarity tasks, including information retrieval, clustering, and classification. General-purpos

model-releasesarxiv-cs-cl
29 Apr 2026
Safety

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes …

DGX agent

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes out. This new research distills the entire debate into a sin

safetydair-ai--x
29 Apr 2026
Safety

Learning-Based Dynamics Modeling and Robust Control for Tendon-Driven Continuum Robots

DGX agent

arXiv:2604.25691v1 Announce Type: new Abstract: Tendon-Driven Continuum Robots (TDCRs) pose significant modeling and control challenges due to complex nonlinearities, such as frictional hysteresis and

safetyarxiv-cs-ro
29 Apr 2026
Research

Named Entity Recognition of Historical Texts via Large Language Model

DGX agent

arXiv:2508.18090v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated remarkable versatility across a wide range of natural language processing tasks and domains. On

researcharxiv-cs-cl
29 Apr 2026
Research

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models

DGX agent

arXiv:2604.25642v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in visual-textual understanding, yet their reliability is critically undermined b

researcharxiv-cs-cv
29 Apr 2026
Research

Privileged Foresight Distillation: Zero-Cost Future Correction for World Action Models

DGX agent

arXiv:2604.25859v1 Announce Type: new Abstract: World action models jointly predict future video and action during training, raising an open question about what role the future-prediction branch actua

researcharxiv-cs-ro
29 Apr 2026
Safety

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

DGX agent

arXiv:2604.25636v1 Announce Type: new Abstract: Unified multimodal models (UMMs) integrate visual understanding and generation within a single framework. For text-to-image (T2I) tasks, this unified ca

safetyarxiv-cs-cv
29 Apr 2026
Research

Rethinking Layer Redundancy in Large Language Models: Calibration Objectives and Search for Depth Pruning

DGX agent

arXiv:2604.24938v1 Announce Type: cross Abstract: Depth pruning improves the inference efficiency of large language models by removing Transformer blocks. Prior work has focused on importance criteria

researcharxiv-cs-cl
29 Apr 2026
Safety

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

DGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

safetyarxiv-cs-cl
29 Apr 2026
Research

Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models

DGX agent

arXiv:2604.25570v1 Announce Type: new Abstract: Spiking Transformers have shown strong potential for long-range visual modeling through spike-driven self-attention. However, their quadratic token inte

researcharxiv-cs-cv
29 Apr 2026
Model Releases

A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models

DGX agent

arXiv:2512.03915v3 Announce Type: replace-cross Abstract: In large-scale AI training, Sparse Mixture-of-Experts (s-MoE) layers enable scaling by activating only a small subset of experts per token. An

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model

DGX agent

arXiv:2508.06206v4 Announce Type: replace-cross Abstract: Affordance grounding focuses on predicting the specific regions of objects that are associated with the actions to be performed by robots. It

model-releasesarxiv-cs-cv
28 Apr 2026
Industry

AI researchers launch talkie, a 13B vintage language model trained on historical text with a 1930 cutoff, to see if it can replicate scientific breakthroughs (talkie)

DGX agent

talkie: AI researchers launch talkie, a 13B vintage language model trained on historical text with a 1930 cutoff, to see if it can replicate scientific breakthroughs — Why vintage language models? — H

industrytechmeme
28 Apr 2026
Local Ai

Breaking the Resource Wall: Geometry-Guided Sequence Modeling for Efficient Semantic Segmentation

DGX agent

arXiv:2604.23399v1 Announce Type: new Abstract: High-performance semantic segmentation has achieved significant progress in recent years, often driven by increasingly large backbones and higher comput

local-aiarxiv-cs-cv
28 Apr 2026
Research

Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation

DGX agent

arXiv:2604.24361v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance in general machine translation, yet their ability in culture-aware scenarios remains poorl

researcharxiv-cs-cl
28 Apr 2026
Tutorials

DeepCausalMMM: A Deep Learning Framework for Marketing Mix Modeling with Causal Structure Learning

DGX agent

arXiv:2510.13087v3 Announce Type: replace Abstract: Marketing Mix Modeling (MMM) estimates the impact of marketing activities on business outcomes such as sales or revenue. Traditional MMM approaches

tutorialsarxiv-cs-lg
28 Apr 2026
Applications

DualGuard: Dual-stream Large Language Model Watermarking Defense against Paraphrase and Spoofing Attack

DGX agent

arXiv:2512.16182v2 Announce Type: replace-cross Abstract: With the rapid development of cloud-based services, large language models have become increasingly accessible through various web platforms. H

applicationsarxiv-cs-cl
28 Apr 2026
Research

Emotion-Conditioned Short-Horizon Human Pose Forecasting with a Lightweight Predictive World Model

DGX agent

arXiv:2604.23532v1 Announce Type: cross Abstract: Short-term human pose prediction plays a crucial role in interactive systems, assistive robots, and emotion-aware human-computer interaction[1-3]. Whi

researcharxiv-cs-ai
28 Apr 2026
Model Releases

FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment

DGX agent

arXiv:2604.23786v1 Announce Type: new Abstract: In recent years, the integration of multimodal machine learning in wellbeing assessment has offered transformative potential for monitoring mental healt

model-releasesarxiv-cs-ai
28 Apr 2026
Hardware

FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost

DGX agent

arXiv:2604.24073v1 Announce Type: cross Abstract: Modern industrial Deep Learning Recommendation Models typically extract user preferences through the analysis of sequential interaction histories, sub

hardwarearxiv-cs-ai
28 Apr 2026
Local Ai

Local model users get a lot of boring-good fixes: @ollama context handling, thinking controls, timeouts, local auth, discovery, and OpenAI-c…

DGX agent

Local model users get a lot of boring-good fixes: @ollama context handling, thinking controls, timeouts, local auth, discovery, and OpenAI-compatible proxy behavior. https://docs.openclaw.ai/providers

local-aiollama--x
28 Apr 2026
Tutorials

LunarDepthNet: Generation of Digital Elevation Models using Deep Learning and Monocular Satellite Images

DGX agent

arXiv:2604.22848v1 Announce Type: new Abstract: Recent times have seen an increase in demand of high quality Digital Elevation Models (DEMs) for the lunar surface, because they are highly important fo

tutorialsarxiv-cs-cv
28 Apr 2026
← Previous
1…175176177178179…1263
Next →