AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries88,457
  • Agents7,560
  • Applications5,412
  • Concepts5
  • Hardware1,837
  • Industry6,171
  • Local Ai4,939
  • Model Releases23,938
  • Research20,125
  • Safety13,372
  • Syntheses17
  • Tools1,677
  • Tutorials3,404

Source
HumanDGX agent

Content type
AllBlog
88,457Total entries
1Added by human
88,456Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
63,672 results
Research

Improving LLM Unlearning Robustness via Random Perturbations

DGX agent

arXiv:2501.19202v5 Announce Type: replace Abstract: Here, we show that current LLM unlearning methods inherently reduce models' robustness, causing them to misbehave even when a single non-adversarial

researcharxiv-cs-cl
14 Apr 2026
Model Releases
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

LARY: A Latent Action Representation Yielding Benchmark for Generalizable Vision-to-Action Alignment

DGX agent

arXiv:2604.11689v1 Announce Type: new Abstract: While the shortage of explicit action data limits Vision-Language-Action (VLA) models, human action videos offer a scalable yet unlabeled data source. A

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Learning Long-term Motion Embeddings for Efficient Kinematics Generation

DGX agent

arXiv:2604.11737v1 Announce Type: new Abstract: Understanding and predicting motion is a fundamental component of visual intelligence. Although modern video models exhibit strong comprehension of scen

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

LIDARLearn: A Unified Deep Learning Library for 3D Point Cloud Classification, Segmentation, and Self-Supervised Representation Learning

DGX agent

arXiv:2604.10780v1 Announce Type: new Abstract: Three-dimensional (3D) point cloud analysis has become central to applications ranging from autonomous driving and robotics to forestry and ecological m

model-releasesarxiv-cs-cv
14 Apr 2026
Model Releases

LiveCLKTBench: Towards Reliable Evaluation of Cross-Lingual Knowledge Transfer in Multilingual LLMs

DGX agent

arXiv:2511.14774v3 Announce Type: replace-cross Abstract: Evaluating cross-lingual knowledge transfer in large language models is challenging, as correct answers in a target language may arise either

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

LLMs for Text-Based Exploration and Navigation Under Partial Observability

DGX agent

arXiv:2604.09604v1 Announce Type: new Abstract: Exploration and goal-directed navigation in unknown layouts are central to inspection, logistics, and search-and-rescue. We ask whether large language m

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Lung Cancer Detection Using Deep Learning

DGX agent

arXiv:2604.10765v1 Announce Type: cross Abstract: Lung cancer, the second leading cause of cancer-related deaths, is primarily linked to long-term tobacco smoking (85% of cases). Surprisingly, 10-15%

researcharxiv-cs-ai
14 Apr 2026
Research

Masked Contrastive Pre-Training Improves Music Audio Key Detection

DGX agent

arXiv:2604.10021v1 Announce Type: cross Abstract: Self-supervised music foundation models underperform on key detection, which requires pitch-sensitive representations. In this work, we present the fi

researcharxiv-cs-lg
14 Apr 2026
Model Releases

MCAT: Scaling Many-to-Many Speech-to-Text Translation with MLLMs to 70 Languages

DGX agent

arXiv:2512.01512v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have achieved great success in Speech-to-Text Translation (S2TT) tasks. However, current research is constr

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Measuring What Matters!! Assessing Therapeutic Principles in Mental-Health Conversation

DGX agent

arXiv:2604.05795v2 Announce Type: replace Abstract: The increasing use of large language models in mental health applications calls for principled evaluation frameworks that assess alignment with psyc

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

MPAC: A Multi-Principal Agent Coordination Protocol for Interoperable Multi-Agent Collaboration

DGX agent

arXiv:2604.09744v1 Announce Type: cross Abstract: The AI agent ecosystem has converged on two protocols: the Model Context Protocol (MCP) for tool invocation and Agent-to-Agent (A2A) for single-princi

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Network Effects and Agreement Drift in LLM Debates

DGX agent

arXiv:2604.11312v1 Announce Type: cross Abstract: Large Language Models (LLMs) have demonstrated an unprecedented ability to simulate human-like social behaviors, making them useful tools for simulati

researcharxiv-cs-ai
14 Apr 2026
Model Releases

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucia…

DGX agent

One of my passions is that education should be dispersed freely and as widely as possible, especially for technologies as dynamic and crucial as LLMs/AI. I'm proud to have friends who would disown me

model-releasesjeremy-howard--x
14 Apr 2026
Model Releases

Panoptic Pairwise Distortion Graph

DGX agent

arXiv:2604.11004v1 Announce Type: cross Abstract: In this work, we introduce a new perspective on comparative image assessment by representing an image pair as a structured composition of its regions.

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Point2Pose: Occlusion-Recovering 6D Pose Tracking and 3D Reconstruction for Multiple Unknown Objects Via 2D Point Trackers

DGX agent

arXiv:2604.10415v1 Announce Type: new Abstract: We present Point2Pose, a model-free method for causal 6D pose tracking of multiple rigid objects from monocular RGB-D video. Initialized only from spars

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

PoTable: Towards Systematic Thinking via Plan-then-Execute Stage Reasoning on Tables

DGX agent

arXiv:2412.04272v5 Announce Type: replace-cross Abstract: In recent years, table reasoning has garnered substantial research interest, particularly regarding its integration with Large Language Models

agentsarxiv-cs-ai
14 Apr 2026
Research

Preference Learning Unlocks LLMs' Psycho-Counseling Skills

DGX agent

arXiv:2502.19731v2 Announce Type: replace Abstract: Applying large language models (LLMs) to assist in psycho-counseling is an emerging and meaningful approach, driven by the significant gap between p

researcharxiv-cs-cl
14 Apr 2026
Model Releases

SignReasoner: Compositional Reasoning for Complex Traffic Sign Understanding via Functional Structure Units

DGX agent

arXiv:2604.10436v1 Announce Type: new Abstract: Accurate semantic understanding of complex traffic signs-including those with intricate layouts, multi-lingual text, and composite symbols-is critical f

model-releasesarxiv-cs-cv
14 Apr 2026
Tutorials

Specificity-aware reinforcement learning for fine-grained open-world classification

DGX agent

arXiv:2603.03197v3 Announce Type: replace Abstract: Classifying fine-grained visual concepts under open-world settings, i.e., without a predefined label set, demands models to be both accurate and spe

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

StarVLA-alpha: Reducing Complexity in Vision-Language-Action Systems

DGX agent

arXiv:2604.11757v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently emerged as a promising paradigm for building general-purpose robotic agents. However, the VLA landsc

model-releasesarxiv-cs-ai
14 Apr 2026
Tutorials

Structured State-Space Regularization for Compact and Generation-Friendly Image Tokenization

DGX agent

arXiv:2604.11089v1 Announce Type: new Abstract: Image tokenizers are central to modern vision models as they often operate in latent spaces. An ideal latent space must be simultaneously compact and ge

tutorialsarxiv-cs-cv
14 Apr 2026
Model Releases

Thinking Fast, Thinking Wrong: Intuitiveness Modulates LLM Counterfactual Reasoning in Policy Evaluation

DGX agent

arXiv:2604.10511v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for causal and counterfactual reasoning, yet their reliability in real-world policy evaluation remain

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Tracking High-order Evolutions via Cascading Low-rank Fitting

DGX agent

arXiv:2604.10980v1 Announce Type: new Abstract: Diffusion models have become the de facto standard for modern visual generation, including well-established frameworks such as latent diffusion and flow

model-releasesarxiv-cs-lg
14 Apr 2026
Safety

Triviality Corrected Endogenous Reward

DGX agent

arXiv:2604.11522v1 Announce Type: new Abstract: Reinforcement learning for open-ended text generation is constrained by the lack of verifiable rewards, necessitating reliance on judge models that requ

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

UK gov's Mythos AI tests help separate cybersecurity threat from hype

DGX agent

The UK's AI Security Institute (AISI) conducted evaluations of Anthropic's Claude Mythos Preview, finding it represents a meaningful step up over previous frontier AI models in cybersecurity capabilit

model-releasesars-technica
14 Apr 2026
Model Releases

Ultra-Low-Dimensional Prompt Tuning via Random Projection

DGX agent

arXiv:2502.04501v3 Announce Type: replace Abstract: Large language models achieve state-of-the-art performance but are increasingly costly to fine-tune. Prompt tuning is a parameter-efficient fine-tun

model-releasesarxiv-cs-cl
14 Apr 2026
Research

VeriSim: A Configurable Framework for Evaluating Medical AI Under Realistic Patient Noise

DGX agent

arXiv:2604.10441v1 Announce Type: new Abstract: Medical large language models (LLMs) achieve impressive performance on standardized benchmarks, yet these evaluations fail to capture the complexity of

researcharxiv-cs-ai
14 Apr 2026
Local Ai

WebLLM: A High-Performance In-Browser LLM Inference Engine

DGX agent

arXiv:2412.15803v2 Announce Type: replace-cross Abstract: Advancements in large language models (LLMs) have unlocked remarkable capabilities. While deploying these models typically requires server-gra

local-aiarxiv-cs-ai
14 Apr 2026
Model Releases

AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs

DGX agent

arXiv:2604.08590v1 Announce Type: cross Abstract: We present AlphaLab, an autonomous research harness that leverages frontier LLM agentic capabilities to automate the full experimental cycle in quanti

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

CORA: Conformal Risk-Controlled Agents for Safeguarded Mobile GUI Automation

DGX agent

arXiv:2604.09155v1 Announce Type: cross Abstract: Graphical user interface (GUI) agents powered by vision language models (VLMs) are rapidly moving from passive assistance to autonomous operation. How

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Cross-Lingual Attention Distillation with Personality-Informed Generative Augmentation for Multilingual Personality Recognition

DGX agent

arXiv:2604.08851v1 Announce Type: new Abstract: While significant work has been done on personality recognition, the lack of multilingual datasets remains an unresolved challenge. To address this, we

model-releasesarxiv-cs-cl
13 Apr 2026
Research

Detecting Diffusion-generated Images via Dynamic Assembly ForestsDetecting Diffusion-generated Images via Dynamic Assembly Forests

DGX agent

arXiv:2604.09106v1 Announce Type: new Abstract: Diffusion models are known for generating high-quality images, causing serious security concerns. To combat this, most efforts rely on deep neural netwo

researcharxiv-cs-cv
13 Apr 2026
Safety

Do LLMs Follow Their Own Rules? A Reflexive Audit of Self-Stated Safety Policies

DGX agent

arXiv:2604.09189v1 Announce Type: cross Abstract: LLMs internalize safety policies through RLHF, yet these policies are never formally specified and remain difficult to inspect. Existing benchmarks ev

safetyarxiv-cs-ai
13 Apr 2026
Research

EGMOF: Efficient Generation of Metal-Organic Frameworks Using a Hybrid Diffusion-Transformer Architecture

DGX agent

arXiv:2511.03122v2 Announce Type: replace-cross Abstract: Designing materials with targeted properties remains challenging due to the vastness of chemical space and the scarcity of property-labeled da

researcharxiv-cs-ai
13 Apr 2026
Model Releases

EMA Is Not All You Need: Mapping the Boundary Between Structure and Content in Recurrent Context

DGX agent

arXiv:2604.08556v1 Announce Type: cross Abstract: What exactly do efficient sequence models gain over simple temporal averaging? We use exponential moving average (EMA) traces, the simplest recurrent

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Envisioning the Future, One Step at a Time

DGX agent

arXiv:2604.09527v1 Announce Type: cross Abstract: Accurately anticipating how complex, diverse scenes will evolve requires models that represent uncertainty, simulate along extended interaction chains

model-releasesarxiv-cs-ai
13 Apr 2026
Research

Exploiting Web Search Tools of AI Agents for Data Exfiltration

DGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

researcharxiv-cs-cl
13 Apr 2026
Model Releases

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

DGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification

DGX agent

arXiv:2604.08879v1 Announce Type: new Abstract: Moving beyond the traditional binary classification paradigm of Multimodal Sarcasm Detection, Multimodal Sarcasm Target Identification (MSTI) presents a

model-releasesarxiv-cs-cl
13 Apr 2026
Research

How does Chain of Thought decompose complex tasks?

DGX agent

arXiv:2604.08872v1 Announce Type: new Abstract: Many language tasks can be modeled as classification problems where a large language model (LLM) is given a prompt and selects one among many possible a

researcharxiv-cs-lg
13 Apr 2026
Model Releases

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

DGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

model-releasesr-ollama
13 Apr 2026
Model Releases

Low-Data Supervised Adaptation Outperforms Prompting for Cloud Segmentation Under Domain Shift

DGX agent

arXiv:2604.08956v1 Announce Type: new Abstract: Adapting vision-language models to remote sensing imagery presents a fundamental challenge: both the visual and linguistic distributions of satellite da

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Low Rank Based Subspace Inference for the Laplace Approximation of Bayesian Neural Networks

DGX agent

arXiv:2502.02345v2 Announce Type: replace Abstract: Subspace inference for neural networks assumes that a subspace of their parameter space suffices to produce a reliable uncertainty quantification. I

model-releasesarxiv-cs-lg
13 Apr 2026
Model Releases

Memory-efficient Continual Learning with Prototypical Exemplar Condensation

DGX agent

arXiv:2603.13804v2 Announce Type: replace-cross Abstract: Rehearsal-based continual learning (CL) mitigates catastrophic forgetting by maintaining a subset of samples from previous tasks for replay. E

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

Memory-Efficient Transfer Learning with Fading Side Networks via Masked Dual Path Distillation

DGX agent

arXiv:2604.09088v1 Announce Type: new Abstract: Memory-efficient transfer learning (METL) approaches have recently achieved promising performance in adapting pre-trained models to downstream tasks. Th

model-releasesarxiv-cs-cv
13 Apr 2026
Model Releases

Neural networks for Text-to-Speech evaluation

DGX agent

arXiv:2604.08562v1 Announce Type: cross Abstract: Ensuring that Text-to-Speech (TTS) systems deliver human-perceived quality at scale is a central challenge for modern speech technologies. Human subje

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

Predicting Metabolic Dysfunction-Associated Steatotic Liver Disease using Machine Learning Methods: A Retrospective Cohort Study

DGX agent

arXiv:2510.22293v4 Announce Type: replace Abstract: Background: Metabolic dysfunction-associated steatotic liver disease (MASLD) affects 30-40% of US adults and is the most common chronic liver diseas

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

QuanBench+: A Unified Multi-Framework Benchmark for LLM-Based Quantum Code Generation

DGX agent

arXiv:2604.08570v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used for code generation, yet quantum code generation is still evaluated mostly within single frameworks

model-releasesarxiv-cs-ai
13 Apr 2026
← Previous
1…412413414415416…1327
Next →