AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,537 results
30 Apr 2026

Sociodemographic Biases in Educational Counselling by Large Language Models

SafetyDGX agent

arXiv:2604.25932v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into educational settings, understanding their potential biases is critical. This study ex

STARRY: Spatial-Temporal Action-Centric World Modeling for Robotic Manipulation

SafetyDGX agent

arXiv:2604.26848v1 Announce Type: new Abstract: Robotic manipulation critically requires reasoning about future spatial-temporal interactions, yet existing VLA policies and world-model-enhanced polici

The Tesla Model X was the fastest-selling used vehicle in the U.S. in March 2026, according to a new study by iSeeCars. On average, it found…

IndustryDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The Tesla Model X was the fastest-selling used vehicle in the U.S. in March 2026, according to a new study by iSeeCars. On average, it found a buyer in 26 days, compared to the typical car, which usua

29 Apr 2026

A multi-stage soft computing framework for complex disease modelling and decision support: A liver cirrhosis case study

ApplicationsDGX agent

arXiv:2604.24796v1 Announce Type: cross Abstract: Liver cirrhosis is a major global health problem causing millions of deaths annually, and timely detection with aggressive treatment can significantly

ADE: Adaptive Dictionary Embeddings -- Scaling Multi-Anchor Representations to Large Language Models

Model ReleasesDGX agent

arXiv:2604.24940v1 Announce Type: new Abstract: Word embeddings are fundamental to natural language processing, yet traditional approaches represent each word with a single vector, creating representa

After switching into the new directory (`cd lc-docs`) I'm going to set the name of my agent, and the model to use This is done in deepagents…

AgentsDGX agent

After switching into the new directory (`cd lc-docs`) I'm going to set the name of my agent, and the model to use This is done in deepagents.toml I'm going to use @Zai_org GLM5 model, served via @base

jina-embeddings-v5-text: Task-Targeted Embedding Distillation

Model ReleasesDGX agent

arXiv:2602.15547v2 Announce Type: replace Abstract: Text embedding models are widely used for semantic similarity tasks, including information retrieval, clustering, and classification. General-purpos

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes …

SafetyDGX agent

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes out. This new research distills the entire debate into a sin

Learning-Based Dynamics Modeling and Robust Control for Tendon-Driven Continuum Robots

SafetyDGX agent

arXiv:2604.25691v1 Announce Type: new Abstract: Tendon-Driven Continuum Robots (TDCRs) pose significant modeling and control challenges due to complex nonlinearities, such as frictional hysteresis and

Named Entity Recognition of Historical Texts via Large Language Model

ResearchDGX agent

arXiv:2508.18090v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have demonstrated remarkable versatility across a wide range of natural language processing tasks and domains. On

Prefill-Time Intervention for Mitigating Hallucination in Large Vision-Language Models

ResearchDGX agent

arXiv:2604.25642v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have achieved remarkable progress in visual-textual understanding, yet their reliability is critically undermined b

Privileged Foresight Distillation: Zero-Cost Future Correction for World Action Models

ResearchDGX agent

arXiv:2604.25859v1 Announce Type: new Abstract: World action models jointly predict future video and action during training, raising an open question about what role the future-prediction branch actua

Refinement via Regeneration: Enlarging Modification Space Boosts Image Refinement in Unified Multimodal Models

SafetyDGX agent

arXiv:2604.25636v1 Announce Type: new Abstract: Unified multimodal models (UMMs) integrate visual understanding and generation within a single framework. For text-to-image (T2I) tasks, this unified ca

Rethinking Layer Redundancy in Large Language Models: Calibration Objectives and Search for Depth Pruning

ResearchDGX agent

arXiv:2604.24938v1 Announce Type: cross Abstract: Depth pruning improves the inference efficiency of large language models by removing Transformer blocks. Prior work has focused on importance criteria

TouchAI: Exploring human-AI perceptual alignment in touch through language model representations

SafetyDGX agent

arXiv:2406.06587v2 Announce Type: replace Abstract: Aligning large language models (LLMs) behaviour with human intent is critical for future AI. An important yet often overlooked aspect of this alignm

Vision SmolMamba: Spike-Guided Token Pruning for Energy-Efficient Spiking State-Space Vision Models

ResearchDGX agent

arXiv:2604.25570v1 Announce Type: new Abstract: Spiking Transformers have shown strong potential for long-range visual modeling through spike-driven self-attention. However, their quadratic token inte

28 Apr 2026

A Theoretical Framework for Auxiliary-Loss-Free Load Balancing of Sparse Mixture-of-Experts in Large-Scale AI Models

Model ReleasesDGX agent

arXiv:2512.03915v3 Announce Type: replace-cross Abstract: In large-scale AI training, Sparse Mixture-of-Experts (s-MoE) layers enable scaling by activating only a small subset of experts per token. An

Affordance-R1: Reinforcement Learning for Generalizable Affordance Reasoning in Multimodal Large Language Model

Model ReleasesDGX agent

arXiv:2508.06206v4 Announce Type: replace-cross Abstract: Affordance grounding focuses on predicting the specific regions of objects that are associated with the actions to be performed by robots. It

AI researchers launch talkie, a 13B vintage language model trained on historical text with a 1930 cutoff, to see if it can replicate scientific breakthroughs (talkie)

IndustryDGX agent

talkie: AI researchers launch talkie, a 13B vintage language model trained on historical text with a 1930 cutoff, to see if it can replicate scientific breakthroughs — Why vintage language models? — H

Breaking the Resource Wall: Geometry-Guided Sequence Modeling for Efficient Semantic Segmentation

Local AiDGX agent

arXiv:2604.23399v1 Announce Type: new Abstract: High-performance semantic segmentation has achieved significant progress in recent years, often driven by increasingly large backbones and higher comput

Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation

ResearchDGX agent

arXiv:2604.24361v1 Announce Type: new Abstract: Large language models (LLMs) have achieved strong performance in general machine translation, yet their ability in culture-aware scenarios remains poorl

DeepCausalMMM: A Deep Learning Framework for Marketing Mix Modeling with Causal Structure Learning

TutorialsDGX agent

arXiv:2510.13087v3 Announce Type: replace Abstract: Marketing Mix Modeling (MMM) estimates the impact of marketing activities on business outcomes such as sales or revenue. Traditional MMM approaches

DualGuard: Dual-stream Large Language Model Watermarking Defense against Paraphrase and Spoofing Attack

ApplicationsDGX agent

arXiv:2512.16182v2 Announce Type: replace-cross Abstract: With the rapid development of cloud-based services, large language models have become increasingly accessible through various web platforms. H

Emotion-Conditioned Short-Horizon Human Pose Forecasting with a Lightweight Predictive World Model

ResearchDGX agent

arXiv:2604.23532v1 Announce Type: cross Abstract: Short-term human pose prediction plays a crucial role in interactive systems, assistive robots, and emotion-aware human-computer interaction[1-3]. Whi

FAIR_XAI: Improving Multimodal Foundation Model Fairness via Explainability for Wellbeing Assessment

Model ReleasesDGX agent

arXiv:2604.23786v1 Announce Type: new Abstract: In recent years, the integration of multimodal machine learning in wellbeing assessment has offered transformative potential for monitoring mental healt

FreeScale: Distributed Training for Sequence Recommendation Models with Minimal Scaling Cost

HardwareDGX agent

arXiv:2604.24073v1 Announce Type: cross Abstract: Modern industrial Deep Learning Recommendation Models typically extract user preferences through the analysis of sequential interaction histories, sub

Local model users get a lot of boring-good fixes: @ollama context handling, thinking controls, timeouts, local auth, discovery, and OpenAI-c…

Local AiDGX agent

Local model users get a lot of boring-good fixes: @ollama context handling, thinking controls, timeouts, local auth, discovery, and OpenAI-compatible proxy behavior. https://docs.openclaw.ai/providers

LunarDepthNet: Generation of Digital Elevation Models using Deep Learning and Monocular Satellite Images

TutorialsDGX agent

arXiv:2604.22848v1 Announce Type: new Abstract: Recent times have seen an increase in demand of high quality Digital Elevation Models (DEMs) for the lunar surface, because they are highly important fo

MetaGAI: A Large-Scale and High-Quality Benchmark for Generative AI Model and Data Card Generation

Model ReleasesDGX agent

arXiv:2604.23539v1 Announce Type: new Abstract: The rapid proliferation of Generative AI necessitates rigorous documentation standards for transparency and governance. However, manual creation of Mode

Modeling Induced Pleasure through Cognitive Appraisal Prediction via Multimodal Fusion

ResearchDGX agent

arXiv:2604.23753v1 Announce Type: new Abstract: Multimodal affective computing analyzes user-generated social media content to predict emotional states. However, a critical gap remains in understandin

Modular Sensory Stream for Integrating Physical Feedback in Vision-Language-Action Models

ApplicationsDGX agent

arXiv:2604.23272v1 Announce Type: new Abstract: Humans understand and interact with the real world by relying on diverse physical feedback beyond visual perception. Motivated by this, recent approache

OntoLogX: Ontology-Guided Knowledge Graph Extraction from Cybersecurity Logs with Large Language Models

Model ReleasesDGX agent

arXiv:2510.01409v2 Announce Type: replace Abstract: System logs represent a valuable source of Cyber Threat Intelligence (CTI), capturing attacker behaviors, exploited vulnerabilities, and traces of m

Reflective Flow Sampling Enhancement

SafetyDGX agent

arXiv:2603.06165v2 Announce Type: replace-cross Abstract: The growing demand for text-to-image generation has led to rapid advances in generative modeling. Recently, text-to-image diffusion models tra

RL Token: Bootstrapping Online RL with Vision-Language-Action Models

SafetyDGX agent

arXiv:2604.23073v1 Announce Type: new Abstract: Vision-language-action (VLA) models can learn to perform diverse manipulation skills 'out of the box,' but achieving the precision and speed that real-w

SCRIBE: Structured Mid-Level Supervision for Tool-Using Language Models

AgentsDGX agent

arXiv:2601.03555v2 Announce Type: replace Abstract: Training reliable tool-augmented agents remains a significant challenge, largely due to the difficulty of credit assignment in multi-step reasoning.

Secure On-Premise Deployment of Open-Weights Large Language Models in Radiology: An Isolation-First Architecture with Prospective Pilot Evaluation

Model ReleasesDGX agent

arXiv:2604.22768v1 Announce Type: cross Abstract: Purpose: To design, implement, evaluate, and report on the regulatory requirements of a self-hosted LLM infrastructure for radiology adhering to the p

Seer: Language Instructed Video Prediction with Latent Diffusion Models

SafetyDGX agent

arXiv:2303.14897v4 Announce Type: replace Abstract: Imagining the future trajectory is the key for robots to make sound planning and successfully reach their goals. Therefore, text-conditioned video p

Stress-Testing Emotional Support Models: Moving from Homogeneous to Diverse Help Seekers

Model ReleasesDGX agent

arXiv:2601.07698v2 Announce Type: replace Abstract: As emotional support chatbots have recently gained significant traction across both research and industry, a common evaluation strategy has emerged:

TexOCR: Advancing Document OCR Models for Compilable Page-to-LaTeX Reconstruction

Model ReleasesDGX agent

arXiv:2604.22880v1 Announce Type: new Abstract: Existing document OCR largely targets plain text or Markdown, discarding the structural and executable properties that make LaTeX essential for scientif

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

AgentsDGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

The Surprising Effectiveness of Membership Inference with Simple N-Gram Coverage

Model ReleasesDGX agent

arXiv:2508.09603v2 Announce Type: replace Abstract: Membership inference attacks serves as useful tool for fair use of language models, such as detecting potential copyright infringement and auditing

Training Machine Learning Models on Encrypted Data: A Privacy-Preserving Framework using Homomorphic Encryption

ApplicationsDGX agent

arXiv:2604.23245v1 Announce Type: cross Abstract: The use of Machine Learning (ML) for data-driven decision-making often relies on access to sensitive datasets, which introduces privacy challenges. Tr

Variational Grey-Box Dynamics Matching

ApplicationsDGX agent

arXiv:2602.17477v3 Announce Type: replace Abstract: Deep generative models such as flow matching and diffusion models have shown great potential in learning complex distributions and dynamical systems

VS-DDPM: Efficient Low-Cost Diffusion Model for Medical Modality Translation

ResearchDGX agent

arXiv:2604.22942v1 Announce Type: cross Abstract: Diffusion models produce high-quality synthetic data but suffer from slow inference. We propose 3D Variable-Step Denoising Diffusion Probabilistic Mod

Weakly Supervised Multicenter Nancy Index Scoring in Ulcerative Colitis Using Foundation Models

ResearchDGX agent

arXiv:2604.23706v1 Announce Type: new Abstract: Histologic assessment of ulcerative colitis (UC) activity is an important endpoint in clinical trials and routine care, but manual grading with indices

When Silence Matters: The Impact of Irrelevant Audio on Text Reasoning in Large Audio-Language Models

ApplicationsDGX agent

arXiv:2510.00626v3 Announce Type: replace-cross Abstract: Large audio-language models (LALMs) unify speech and text processing, but their robustness in noisy real-world settings remains underexplored.

27 Apr 2026

Clutter-Robust Vision-Language-Action Models through Object-Centric and Geometry Grounding

SafetyDGX agent

arXiv:2512.22519v2 Announce Type: replace Abstract: Recent Vision-Language-Action (VLA) models have made impressive progress toward general-purpose robotic manipulation by post-training large Vision-L

DVGT-2: Vision-Geometry-Action Model for Autonomous Driving at Scale

AgentsDGX agent

arXiv:2604.00813v3 Announce Type: replace-cross Abstract: End-to-end autonomous driving has evolved from the conventional paradigm based on sparse perception into vision-language-action (VLA) models,

Initial results of the Digital Consciousness Model

ResearchDGX agent

arXiv:2601.17060v2 Announce Type: replace-cross Abstract: Artificially intelligent systems have become remarkably sophisticated. They hold conversations, write essays, and seem to understand context i

LLMPhy: Parameter-Identifiable Physical Reasoning Combining Large Language Models and Physics Engines

Model ReleasesDGX agent

arXiv:2411.08027v3 Announce Type: replace-cross Abstract: Most learning-based approaches to complex physical reasoning sidestep the crucial problem of parameter identification (e.g., mass, friction) t

Preference Heads in Large Language Models: A Mechanistic Framework for Interpretable Personalization

ResearchDGX agent

arXiv:2604.22345v1 Announce Type: new Abstract: Large Language Models (LLMs) exhibit strong implicit personalization ability, yet most existing approaches treat this behavior as a black box, relying o

Shard the Gradient, Scale the Model: Serverless Federated Aggregation via Gradient Partitioning

ResearchDGX agent

arXiv:2604.22072v1 Announce Type: cross Abstract: Federated learning (FL) aggregation on serverless platforms faces a hard scalability ceiling: existing architectures (lambda-FL, LIFL) partition clien

Source-Modality Monitoring in Vision-Language Models

AgentsDGX agent

arXiv:2604.22038v1 Announce Type: new Abstract: We define and investigate source-modality monitoring -- the ability of multimodal models to track and communicate the input source from which pieces of

Sum-of-Checks: Structured Reasoning for Surgical Safety with Large Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.22156v1 Announce Type: cross Abstract: Purpose: Accurate assessment of the Critical View of Safety (CVS) during laparoscopic cholecystectomy is essential to prevent bile duct injury, a comp

Survey Response Generation: Generating Closed-Ended Survey Responses In-Silico with Large Language Models

SafetyDGX agent

arXiv:2510.11586v2 Announce Type: replace Abstract: Many in-silico simulations of human survey responses with large language models (LLMs) focus on generating closed-ended survey responses, whereas LL

that said, another tweet by the same user does correctly assess what’s wrong with most user’s model of what LLMs say.

SafetyDGX agent

Gary Marcus critiques common misconceptions about how large language models (LLMs) function, suggesting that most users have an incorrect mental model of what LLMs actually do when generating response

Transferable Physical-World Adversarial Patches Against Pedestrian Detection Models

SafetyDGX agent

arXiv:2604.22552v1 Announce Type: new Abstract: Physical adversarial patch attacks critically threaten pedestrian detection, causing surveillance and autonomous driving systems to miss pedestrians and

UniSonate: A Unified Model for Speech, Music, and Sound Effect Generation with Text Instructions

ResearchDGX agent

arXiv:2604.22209v1 Announce Type: cross Abstract: Generative audio modeling has largely been fragmented into specialized tasks, text-to-speech (TTS), text-to-music (TTM), and text-to-audio (TTA), each

26 Apr 2026

America needs to go much harder on open source models

IndustryDGX agent

Clem Delangue argues that the United States should increase investment and policy support for open source AI models to maintain competitive advantage and reduce dependence on proprietary systems contr

25 Apr 2026

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼

AgentsDGX agent

Free API credits to beta testers to coordinate frontier models dynamically. See more below 👇🏼 We’re launching the beta for our new commercial AI product: Sakana Fugu 🐡, a multi-agent orchestration sys

← Previous
1…140141142143144…1009
Next →