AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
59,929 results
21 Apr 2026

Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics

Model ReleasesDGX agent

arXiv:2604.17715v1 Announce Type: cross Abstract: Recent advances in large language models for test case generation have improved branch coverage via prompt-engineered mutations. However, they still l

Residual Diffusion Bridge Model for Image Restoration

ResearchDGX agent

arXiv:2510.23116v3 Announce Type: replace Abstract: Diffusion bridge models establish probabilistic paths between arbitrary paired distributions and exhibit great potential for universal image restora

Revisiting Change VQA in Remote Sensing with Structured and Native Multimodal Qwen Models

Model ReleasesDGX agent

arXiv:2604.18429v1 Announce Type: new Abstract: Change visual question answering (Change VQA) addresses the problem of answering natural-language questions about semantic changes between bi-temporal r

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

S2H-DPO: Hardness-Aware Preference Optimization for Vision-Language Models

Model ReleasesDGX agent

arXiv:2604.18512v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in single-image understanding, yet effective reasoning across multiple images remain

Sky2Ground: A Benchmark for Site Modeling under Varying Altitude

Model ReleasesDGX agent

arXiv:2603.13740v3 Announce Type: replace Abstract: We introduce Sky2Ground, a three-view dataset designed for varying altitude camera localization, correspondence learning, and reconstruction. The da

Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts

Model ReleasesDGX agent

arXiv:2604.16926v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models have shown strong potential for learning generalizable representations from large-scale neural data, yet

When Background Matters: Breaking Medical Vision Language Models by Transferable Attack

ResearchDGX agent

arXiv:2604.17318v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, pos

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints

SafetyDGX agent

arXiv:2604.16916v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is primarily evaluated under open-ended generation, where models can mitigate risk by refusing to respo

When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models

ResearchDGX agent

arXiv:2507.13868v2 Announce Type: replace Abstract: Vision-language models (VLMs) increasingly combine visual and textual information to perform complex tasks. However, conflicts between their interna

20 Apr 2026

AdaVFM: Adaptive Vision Foundation Models for Edge Intelligence via LLM-Guided Execution

Local AiDGX agent

arXiv:2604.15622v1 Announce Type: new Abstract: Language-aligned vision foundation models (VFMs) enable versatile visual understanding for always-on contextual AI, but their deployment on edge devices

ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation

Model ReleasesDGX agent

arXiv:2508.10635v3 Announce Type: replace Abstract: Understanding environmental changes from remote sensing imagery is vital for climate resilience, urban planning, and ecosystem monitoring. Yet, curr

FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models

SafetyDGX agent

arXiv:2604.15488v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit undesirable behaviors, such as safety violations and hallucinations. Although inference-time steering offer

How Hypocritical Is Your LLM judge? Listener-Speaker Asymmetries in the Pragmatic Competence of Large Language Models

ResearchDGX agent

arXiv:2604.15873v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly studied as repositories of linguistic knowledge. In this line of work, models are commonly evaluated both

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

SafetyDGX agent

arXiv:2604.15701v1 Announce Type: new Abstract: The significant computational demands of large language models have increased interest in distilling reasoning abilities into smaller models via Chain-o

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

Model ReleasesDGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This Mo…

Model ReleasesDGX agent

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This MoE beast is cooking on benchmarks: 🧠SWE-bench Verified: 73.4 (

Repurposing 3D Generative Model for Autoregressive Layout Generation

Model ReleasesDGX agent

arXiv:2604.16299v1 Announce Type: new Abstract: We introduce LaviGen, a framework that repurposes 3D generative models for 3D layout generation. Unlike previous methods that infer object layouts from

Tabular foundation models for in-context prediction of molecular properties

ApplicationsDGX agent

arXiv:2604.16123v1 Announce Type: new Abstract: Accurate molecular property prediction is central to drug discovery, catalysis, and process design, yet real-world applications are often limited by sma

19 Apr 2026

I gave two MoE models the same vibe coding challenge Qwen3.6 35B A3B (31.8GB) vs Gemma4 26B A4B (23.3GB) Stack: > Unsloth Q6_K_XL > llama.cp…

Model ReleasesDGX agent

I gave two MoE models the same vibe coding challenge Qwen3.6 35B A3B (31.8GB) vs Gemma4 26B A4B (23.3GB) Stack: > Unsloth Q6_K_XL > llama.cpp > Model-card recommended sampling for each 4 prompts, side

18 Apr 2026

ICYMI from a few weeks back, we compiled our learnings around how to achieve Training-Inference Parity in MoE Models. The Fundamental Issue:…

TutorialsDGX agent

Fireworks AI published a compilation of insights on achieving training-inference parity in Mixture of Experts (MoE) models, addressing fundamental challenges in aligning model behavior between trainin

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite d…

Model ReleasesDGX agent

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite despite requiring 12GB VRAM instead of 24GB 🤯 Runs on a SINGL

17 Apr 2026

Anthropic’s new cybersecurity model could get it back in the government’s good graces

Model ReleasesDGX agent

The Trump administration has spent nearly two months fighting with AI company Anthropic. It's dubbed the company a 'RADICAL LEFT, WOKE COMPANY' full of 'Leftwing nut jobs' and a menace to national sec

Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting

ResearchDGX agent

arXiv:2604.14739v1 Announce Type: new Abstract: Large-scale renewable energy deployment introduces pronounced volatility into the electricity system, turning grid operation into a complex stochastic o

Best Ollama model for n8n workflows (RAG, file handling, reasoning) + hardware requirements?

Local AiDGX agent

Models like Qwen3 and Llama 3.2 are commonly used for n8n RAG workflows , with selection depending on use case requirements. Running Ollama models locally requires at least 16 GB of RAM on your device

CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification

Model ReleasesDGX agent

arXiv:2604.14602v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrad

CobwebTM: Probabilistic Concept Formation for Lifelong and Hierarchical Topic Modeling

Model ReleasesDGX agent

arXiv:2604.14489v1 Announce Type: new Abstract: Topic modeling seeks to uncover latent semantic structure in text corpora with minimal supervision. Neural approaches achieve strong performance but req

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’r…

AgentsDGX agent

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’re building self-improvement infra around agents: - mining Tr

DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines

Model ReleasesDGX agent

arXiv:2604.14314v1 Announce Type: cross Abstract: This manuscript introduces DharmaOCR Full and Lite, a pair of specialized small language models (SSLMs) for structured OCR that jointly optimize trans

Domain-Adaptive Model Merging Across Disconnected Modes

ResearchDGX agent

arXiv:2603.05957v2 Announce Type: replace-cross Abstract: Learning across domains is challenging when data cannot be centralized due to privacy or heterogeneity, which limits the ability to train a si

Exploring the flavor structure of leptons via diffusion models

ResearchDGX agent

arXiv:2503.21432v2 Announce Type: replace-cross Abstract: We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelli

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

SafetyDGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

ResearchDGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models

Model ReleasesDGX agent

arXiv:2604.14591v1 Announce Type: new Abstract: We address the problem of prompt-guided image editing in visual autoregressive models. Given a source image and a target text prompt, we aim to modify t

ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking

Model ReleasesDGX agent

arXiv:2506.03487v3 Announce Type: replace-cross Abstract: Reranking is fundamental to information retrieval and retrieval-augmented generation, with recent Large Language Models (LLMs) significantly a

Psychological Steering of Large Language Models

ResearchDGX agent

arXiv:2604.14463v1 Announce Type: new Abstract: Large language models (LLMs) emulate a consistent human-like behavior that can be shaped through activation-level interventions. This paradigm is conver

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

Model ReleasesDGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models

ResearchDGX agent

arXiv:2512.23578v3 Announce Type: replace Abstract: In this paper, we show that when spoken language models (SLMs) are instructed to speak in a specific speaking style at the beginning of a multi-turn

16 Apr 2026

Claude Opus 4.7 is now the default orchestration model powering Computer. It's also available for Max subscribers on Perplexity web, iOS, an…

Model ReleasesDGX agent

Claude Opus 4.7 has been set as the default orchestration model for Anthropic's Computer product. The model is also available to Max subscribers on Perplexity's web platform and iOS application.

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

Model ReleasesDGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

Google’s DeepMind just released new 4B and 27B MedGemma models!

Model ReleasesDGX agent

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification

Model ReleasesDGX agent

arXiv:2604.13285v1 Announce Type: new Abstract: Clinical text classification requires choosing between specialized fine-tuned models (BERT variants) and general-purpose large language models (LLMs), y

Model 3 & Y are the most cost-effective vehicles you can own

IndustryDGX agent

Model 3 & Y are the most cost-effective vehicles you can own This is a great graphic that compares cost per kilometer in Australia between the top 10 ICE vehicles & top 10 EVs. The difference is huge.

ReproMIA: A Comprehensive Analysis of Model Reprogramming for Proactive Membership Inference Attacks

ResearchDGX agent

arXiv:2603.28942v3 Announce Type: replace Abstract: The pervasive deployment of deep learning models across critical domains has concurrently intensified privacy concerns due to their inherent propens

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

Model ReleasesDGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

Simplicity Prevails: The Emergence of Generalizable AIGI Detection in Visual Foundation Models

Local AiDGX agent

arXiv:2602.01738v2 Announce Type: replace Abstract: While specialized detectors for AI-Generated Images (AIGI) achieve near-perfect accuracy on curated benchmarks, they suffer from a dramatic performa

Training and Finetuning Multimodal Embedding & Reranker Models with Sentence Transformers

ToolsDGX agent

This guide covers how to train and fine-tune multimodal embedding and reranker models using the Sentence Transformers library, enabling systems to work with both text and image data simultaneously. It

UNBOX: Unveiling Black-box visual models with Natural-language

SafetyDGX agent

arXiv:2603.08639v2 Announce Type: replace Abstract: Ensuring trustworthiness in open-world visual recognition requires models that are interpretable, fair, and robust to distribution shifts. Yet moder

15 Apr 2026

A Foot Resistive Force Model for Legged Locomotion on Muddy Terrains

Model ReleasesDGX agent

arXiv:2604.12006v1 Announce Type: new Abstract: Legged robots face significant challenges in moving and navigating on deformable and highly yielding terrain such as mud. We present a resistive force m

Beyond Output Correctness: Benchmarking and Evaluating Large Language Model Reasoning in Coding Tasks

Model ReleasesDGX agent

arXiv:2604.12379v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on explicit reasoning to solve coding tasks, yet evaluating the quality of this reasoning remains chall

Bipedal-Walking-Dynamics Model on Granular Terrains

Model ReleasesDGX agent

arXiv:2604.11981v1 Announce Type: new Abstract: Bipeds have demonstrated high agility and mobility in unstructured environments such as sand. The yielding of such granular media brings significant sin

Fragile Preferences: A Deep Dive Into Order Effects in Large Language Models

Model ReleasesDGX agent

arXiv:2506.14092v3 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed in decision-support systems for high-stakes domains such as hiring and university admissions,

Fully Homomorphic Encryption on Llama 3 model for privacy preserving LLM inference

Model ReleasesDGX agent

arXiv:2604.12168v1 Announce Type: cross Abstract: The applications of Generative Artificial Intelligence (GenAI) and their intersections with data-driven fields, such as healthcare, finance, transport

Hard Negative Sample-Augmented DPO Post-Training for Small Language Models

Model ReleasesDGX agent

arXiv:2512.19728v2 Announce Type: replace Abstract: Large language models (LLMs) continue to struggle with mathematical reasoning, and common post-training pipelines often reduce each generated soluti

HintMR: Eliciting Stronger Mathematical Reasoning in Small Language Models

Local AiDGX agent

arXiv:2604.12229v1 Announce Type: new Abstract: Small language models (SLMs) often struggle with complex mathematical reasoning due to limited capacity to maintain long chains of intermediate steps an

Mining Large Language Models for Low-Resource Language Data: Comparing Elicitation Strategies for Hausa and Fongbe

Model ReleasesDGX agent

arXiv:2604.12477v1 Announce Type: cross Abstract: Large language models (LLMs) are trained on data contributed by low-resource language communities, yet the linguistic knowledge encoded in these model

PolicyLLM: Towards Excellent Comprehension of Public Policy for Large Language Models

Model ReleasesDGX agent

arXiv:2604.12995v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly integrated into real-world decision-making, including in the domain of public policy. Yet, their ability t

14 Apr 2026

A Tale of Two Temperatures: Simple, Efficient, and Diverse Sampling from Diffusion Language Models

TutorialsDGX agent

arXiv:2604.09921v1 Announce Type: new Abstract: Much work has been done on designing fast and accurate sampling for diffusion language models (dLLMs). However, these efforts have largely focused on th

Assessing the Pedagogical Readiness of Large Language Models as AI Tutors in Low-Resource Contexts: A Case Study of Nepal's K-10 Curriculum

Model ReleasesDGX agent

arXiv:2604.09619v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into educational ecosystems promises to democratize access to personalized tutoring, yet the readiness

bacpipe: a Python package to make bioacoustic deep learning models accessible

ResearchDGX agent

arXiv:2604.11560v1 Announce Type: cross Abstract: 1. Natural sounds have been recorded for millions of hours over the previous decades using passive acoustic monitoring. Improvements in deep learning

CheeseBench: Evaluating Large Language Models on Rodent Behavioral Neuroscience Paradigms

Model ReleasesDGX agent

arXiv:2604.10825v1 Announce Type: new Abstract: We introduce CheeseBench, a benchmark that evaluates large language models (LLMs) on nine classical behavioral neuroscience paradigms (Morris water maze

← Previous
1…5657585960…999
Next →