AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
Local Ai

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models

DGX agent

arXiv:2602.01738v2 Announce Type: replace Abstract: While specialized detectors for AI-Generated Images (AIGI) achieve near-perfect accuracy on curated benchmarks, they suffer from a dramatic performa

local-aiarxiv-cs-cv
16 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Tools

Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers

DGX agent

This guide covers how to train and fine-tune multimodal embedding and reranker models using the Sentence Transformers library, enabling systems to work with both text and image data simultaneously. It

toolshugging-face
16 Apr 2026
Safety

UNBOX: Unveiling Black-box visual models with Natural-language

DGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

A Foot Resistive Force Model for Legged Locomotion on Muddy Terrains

DGX agent

arXiv:2604.12006v1 Announce Type: new Abstract: Legged robots face significant challenges in moving and navigating on deformable and highly yielding terrain such as mud. We present a resistive force m

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks

DGX agent

arXiv:2604.12379v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on explicit reasoning to solve coding tasks, yet evaluating the quality of this reasoning remains chall

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Bipedal-Walking-Dynamics Model on Granular Terrains

DGX agent

arXiv:2604.11981v1 Announce Type: new Abstract: Bipeds have demonstrated high agility and mobility in unstructured environments such as sand. The yielding of such granular media brings significant sin

model-releasesarxiv-cs-ro
15 Apr 2026
Model Releases

Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models

DGX agent

arXiv:2506.14092v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in decision-support systems for high-stakes domains such as hiring and university admissions,

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Fully Homomorphic Encryption on Llama 3 model for privacy preserving LLM inference

DGX agent

arXiv:2604.12168v1 Announce Type: cross Abstract: The applications of Generative Artificial Intelligence (GenAI) and their intersections with data-driven fields, such as healthcare, finance, transport

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Hard Negative Sample-Augmented DPO Post-Training for Small Language Models

DGX agent

arXiv:2512.19728v2 Announce Type: replace Abstract: Large language models (LLMs) continue to struggle with mathematical reasoning, and common post-training pipelines often reduce each generated soluti

model-releasesarxiv-cs-lg
15 Apr 2026
Local Ai

HintMR: Eliciting Stronger Mathematical Reasoning in Small Language Models

DGX agent

arXiv:2604.12229v1 Announce Type: new Abstract: Small language models (SLMs) often struggle with complex mathematical reasoning due to limited capacity to maintain long chains of intermediate steps an

local-aiarxiv-cs-ai
15 Apr 2026
Model Releases

Mining Large Language Models for Low-Resource Language Data: Comparing Elicitation Strategies for Hausa and Fongbe

DGX agent

arXiv:2604.12477v1 Announce Type: cross Abstract: Large language models (LLMs) are trained on data contributed by low-resource language communities, yet the linguistic knowledge encoded in these model

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models

DGX agent

arXiv:2604.12995v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability t

model-releasesarxiv-cs-cl
15 Apr 2026
Tutorials

A Tale of Two Temperatures: Simple, Efficient, and Diverse Sampling from Diffusion Language Models

DGX agent

arXiv:2604.09921v1 Announce Type: new Abstract: Much work has been done on designing fast and accurate sampling for diffusion language models (dLLMs). However, these efforts have largely focused on th

tutorialsarxiv-cs-lg
14 Apr 2026
Model Releases

Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum

DGX agent

arXiv:2604.09619v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into educational ecosystems promises to democratize access to personalized tutoring, yet the readiness

model-releasesarxiv-cs-ai
14 Apr 2026
Research

bacpipe: a Python package to make bioacoustic deep learning models accessible

DGX agent

arXiv:2604.11560v1 Announce Type: cross Abstract: 1. Natural sounds have been recorded for millions of hours over the previous decades using passive acoustic monitoring. Improvements in deep learning

researcharxiv-cs-ai
14 Apr 2026
Model Releases

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

DGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Comparative Analysis of Large Language Models in Healthcare

DGX agent

arXiv:2604.10316v1 Announce Type: new Abstract: Background: Large Language Models (LLMs) are transforming artificial intelligence applications in healthcare due to their ability to understand, generat

model-releasesarxiv-cs-cl
14 Apr 2026
Research

DA-PTQ: Drift-Aware Post-Training Quantization for Efficient Vision-Language-Action Models

DGX agent

arXiv:2604.11572v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) have demonstrated strong potential for embodied AI, yet their deployment on resource-limited robots remains challen

researcharxiv-cs-ro
14 Apr 2026
Model Releases

Enhancing Multimodal Large Language Models for Ancient Chinese Character Evolution Analysis via Glyph-Driven Fine-Tuning

DGX agent

arXiv:2604.11299v1 Announce Type: cross Abstract: In recent years, rapid advances in Multimodal Large Language Models (MLLMs) have increasingly stimulated research on ancient Chinese scripts. As the e

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Environmental Footprint of GenAI Research: Insights from the Moshi Foundation Model

DGX agent

arXiv:2604.11154v1 Announce Type: new Abstract: New multi-modal large language models (MLLMs) are continuously being trained and deployed, following rapid development cycles. This generative AI frenzy

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Evaluating Reliability Gaps in Large Language Model Safety via Repeated Prompt Sampling

DGX agent

arXiv:2604.09606v1 Announce Type: new Abstract: Traditional benchmarks for large language models (LLMs), such as HELM and AIR-BENCH, primarily assess safety risk through breadth-oriented evaluation ac

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

GeoArena: Evaluating Open-World Geographic Reasoning in Large Vision-Language Models

DGX agent

arXiv:2509.04334v4 Announce Type: replace Abstract: Geographic reasoning is a fundamental cognitive capability that requires models to infer plausible locations by synthesizing visual evidence with sp

model-releasesarxiv-cs-cv
14 Apr 2026
Research

GS4City: Hierarchical Semantic Gaussian Splatting via City-Model Priors

DGX agent

arXiv:2604.11401v1 Announce Type: new Abstract: Recent semantic 3D Gaussian Splatting (3DGS) methods primarily rely on 2D foundation models, often yielding ambiguous boundaries and limited support for

researcharxiv-cs-cv
14 Apr 2026
Model Releases

How Robust Are Large Language Models for Clinical Numeracy? An Empirical Study on Numerical Reasoning Abilities in Clinical Contexts

DGX agent

arXiv:2604.11133v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly being explored for clinical question answering and decision support, yet safe deployment critically requir

model-releasesarxiv-cs-cl
14 Apr 2026
Applications

Inferring Dynamic Physical Properties from Video Foundation Models

DGX agent

arXiv:2510.02311v2 Announce Type: replace Abstract: We study the task of predicting dynamic physical properties from videos. More specifically, we consider physical properties that require temporal in

applicationsarxiv-cs-cv
14 Apr 2026
Safety

Influencing Humans to Conform to Preference Models for RLHF

DGX agent

arXiv:2501.06416v3 Announce Type: replace-cross Abstract: Designing a reinforcement learning from human feedback (RLHF) algorithm to approximate a human's unobservable reward function requires assumin

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

Is There Knowledge Left to Extract? Evidence of Fragility in Medically Fine-Tuned Vision-Language Models

DGX agent

arXiv:2604.09841v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly adapted through domain-specific fine-tuning, yet it remains unclear whether this improves reasoning bey

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Jailbreaking the Matrix: Nullspace Steering for Controlled Model Subversion

DGX agent

arXiv:2604.10326v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak attacks -- inputs designed to bypass safety mechanisms and elicit harmful responses -- despite ad

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

LaMI: Augmenting Large Language Models via Late Multi-Image Fusion

DGX agent

arXiv:2406.13621v2 Announce Type: replace Abstract: Commonsense reasoning often requires both textual and visual knowledge, yet Large Language Models (LLMs) trained solely on text lack visual groundin

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Latent Structure of Affective Representations in Large Language Models

DGX agent

arXiv:2604.07382v2 Announce Type: replace-cross Abstract: The geometric structure of latent representations in large language models (LLMs) is an active area of research, driven in part by its implica

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

MASS: Motion-Aware Spatial-Temporal Grounding for Physics Reasoning and Comprehension in Vision-Language Models

DGX agent

arXiv:2511.18373v2 Announce Type: replace Abstract: Vision Language Models (VLMs) perform well on standard video tasks but struggle with physics-related reasoning involving motion dynamics and spatial

model-releasesarxiv-cs-cv
14 Apr 2026
Agents

MiniMax M2.7 is now available in LM Studio. This model excels at agentic tool calling 🛠️ Requires at least ~138GB to run locally https://lm…

DGX agent

MiniMax M2.7 is a large language model now available for local deployment through LM Studio, notable for its strong performance in agentic tool calling tasks. The model requires a substantial minimum

agentslm-studio--x
14 Apr 2026
Research

Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model

DGX agent

arXiv:2505.23606v4 Announce Type: replace-cross Abstract: Unified generation models aim to handle diverse tasks across modalities -- such as text generation, image generation, and vision-language reas

researcharxiv-cs-cv
14 Apr 2026
Model Releases

NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment

DGX agent

arXiv:2604.11543v1 Announce Type: cross Abstract: Novelty is a core requirement in academic publishing and a central focus of peer review, yet the growing volume of submissions has placed increasing p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

On Harnessing Idle Compute at the Edge for Foundation Model Training

DGX agent

arXiv:2512.22142v2 Announce Type: replace-cross Abstract: The foundation-model ecosystem remains highly centralized because training requires immense compute resources and is therefore largely limited

model-releasesarxiv-cs-lg
14 Apr 2026
Model Releases

PSF-Med: Measuring and Explaining Paraphrase Sensitivity in Medical Vision Language Models

DGX agent

arXiv:2602.21428v2 Announce Type: replace Abstract: Medical Vision Language Models (VLMs) can change their answers when clinicians rephrase the same question, a failure mode that threatens deployment

model-releasesarxiv-cs-cv
14 Apr 2026
Safety

Reasoning Resides in Layers: Restoring Temporal Reasoning in Video-Language Models with Layer-Selective Merging

DGX agent

arXiv:2604.11399v1 Announce Type: cross Abstract: Multimodal adaptation equips large language models (LLMs) with perceptual capabilities, but often weakens the reasoning ability inherited from languag

safetyarxiv-cs-cl
14 Apr 2026
Model Releases

Retrieval-Augmented Large Language Models for Evidence-Informed Guidance on Cannabidiol Use in Older Adults

DGX agent

arXiv:2604.09548v1 Announce Type: cross Abstract: Older adults commonly experience chronic conditions such as pain and sleep disturbances and may consider cannabidiol for symptom management. Safe use

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

Seeing Through the Tool: A Controlled Benchmark for Occlusion Robustness in Foundation Segmentation Models

DGX agent

arXiv:2604.11711v1 Announce Type: new Abstract: Occlusion, where target structures are partially hidden by surgical instruments or overlapping tissues, remains a critical yet underexplored challenge f

model-releasesarxiv-cs-cv
14 Apr 2026
Research

ShapShift: Explaining Model Prediction Shifts with Subgroup Conditional Shapley Values

DGX agent

arXiv:2604.11200v1 Announce Type: cross Abstract: Changes in input distribution can induce shifts in the average predictions of machine learning models. Such prediction shifts may impact downstream bu

researcharxiv-cs-ai
14 Apr 2026
Model Releases

SODA: Semi On-Policy Black-Box Distillation for Large Language Models

DGX agent

arXiv:2604.03873v2 Announce Type: replace-cross Abstract: Black-box knowledge distillation for large language models presents a strict trade-off. Simple off-policy methods (e.g., sequence-level knowle

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Text-to-Image Models and Their Representation of People from Different Nationalities Engaging in Activities

DGX agent

arXiv:2504.06313v5 Announce Type: replace Abstract: This paper investigates how popular text-to-image (T2I) models, DALL-E 3 and Gemini 3 Pro Preview, depict people from 206 nationalities when prompte

model-releasesarxiv-cs-cv
14 Apr 2026
Applications

Tuning Language Models for Robust Prediction of Diverse User Behaviors

DGX agent

arXiv:2505.17682v2 Announce Type: replace-cross Abstract: Predicting user behavior is essential for intelligent assistant services, yet deep learning models often struggle to capture long-tailed behav

applicationsarxiv-cs-ai
14 Apr 2026
Model Releases

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cuttin…

DGX agent

We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than payroll) — cutting it by 2-5x would be transformative. Last year, OSS models

model-releasesharrison-chase--x
14 Apr 2026
Model Releases

Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models

DGX agent

arXiv:2604.10079v1 Announce Type: new Abstract: Supervised Fine-Tuning (SFT) is the standard approach for adapting large language models (LLMs) to downstream tasks. However, we observe a persistent fa

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

Breaking Block Boundaries: Anchor-based History-stable Decoding for Diffusion Large Language Models

DGX agent

arXiv:2604.08964v1 Announce Type: new Abstract: Diffusion Large Language Models (dLLMs) have recently become a promising alternative to autoregressive large language models (ARMs). Semi-autoregressive

model-releasesarxiv-cs-cl
13 Apr 2026
Model Releases

Sources: SoftBank, Sony, Honda, and six other Japanese companies launch a new AI company to develop a 1T-parameter foundation model for 'physical AI' by 2030 (Natsuki Yamamoto/Nikkei Asia)

DGX agent

Natsuki Yamamoto / Nikkei Asia: Sources: SoftBank, Sony, Honda, and six other Japanese companies launch a new AI company to develop a 1T-parameter foundation model for “physical AI” by 2030 — TOKYO —

model-releasestechmeme
13 Apr 2026
Model Releases

Where Vision Becomes Text: Locating the OCR Routing Bottleneck in Vision-Language Models

DGX agent

arXiv:2602.22918v2 Announce Type: replace Abstract: Vision-language models (VLMs) can read text from images, but where does this optical character recognition (OCR) information enter the language proc

model-releasesarxiv-cs-cl
13 Apr 2026
← Previous
1…7273747576…1259
Next →