AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,619
  • Agents7,270
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,100
  • Local Ai4,731
  • Model Releases22,595
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,619Total entries
1Added by human
84,618Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,565 results
Safety

SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future

DGX agent

arXiv:2604.16451v1 Announce Type: new Abstract: Recent advances in visual-language models (VLMs) have led to significant improvements in a plethora of complex multimodal tasks like image captioning, r

safetyarxiv-cs-cl
21 Apr 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Research

Test-Time Perturbation Learning with Delayed Feedback for Vision-Language-Action Models

DGX agent

arXiv:2604.18107v1 Announce Type: new Abstract: Vision-Language-Action models (VLAs) achieve remarkable performance in sequential decision-making but remain fragile to subtle environmental shifts, suc

researcharxiv-cs-cv
21 Apr 2026
Agents

The Global Neural World Model: Spatially Grounded Discrete Topologies for Action-Conditioned Planning

DGX agent

arXiv:2604.16585v1 Announce Type: new Abstract: We present the Global Neural World Model (GNWM), a self-stabilizing framework that achieves topological quantization through balanced continuous entropy

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

TLoRA: Task-aware Low Rank Adaptation of Large Language Models

DGX agent

arXiv:2604.18124v1 Announce Type: new Abstract: Low-Rank Adaptation (LoRA) has become a widely adopted parameter-efficient fine-tuning method for large language models, with its effectiveness largely

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Towards Joint Quantization and Token Pruning of Vision-Language Models

DGX agent

arXiv:2604.17320v1 Announce Type: new Abstract: Deploying Vision-Language Models (VLMs) under aggressive low-bit inference remains challenging because inference cost is dominated by the long visual-to

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Unsupervised Discovery of Intermediate Phase Order in the Frustrated J_1-J_2 Heisenberg Model via Prometheus Framework

DGX agent

arXiv:2602.21468v4 Announce Type: replace-cross Abstract: The spin-1/2 J_1-J_2 Heisenberg model on the square lattice exhibits a debated intermediate phase between Neel antiferromagnetic and stripe or

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

When Text Hijacks Vision: Benchmarking and Mitigating Text Overlay-Induced Hallucination in Vision Language Models

DGX agent

arXiv:2604.17375v1 Announce Type: new Abstract: Recent advances in Vision-Language Models (VLMs) have substantially enhanced their ability across multimodal video understanding benchmarks spanning tem

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

AI reviewers then ranked the submissions, and gave the same ordering every time, regardless of model doing the ranking: Codex GPT-5.4 > GPT-…

DGX agent

AI reviewers then ranked the submissions, and gave the same ordering every time, regardless of model doing the ranking: Codex GPT-5.4 > GPT-5.3-Codex > Opus 4.6 > humans. Paper: http://claude-code-eco

model-releasesethan-mollick--x
20 Apr 2026
Applications

Applied Explainability for Large Language Models: A Comparative Study

DGX agent

arXiv:2604.15371v1 Announce Type: cross Abstract: Large language models (LLMs) achieve strong performance across many natural language processing tasks, yet their decision processes remain difficult t

applicationsarxiv-cs-ai
20 Apr 2026
Safety

AutoDrive-R^2: Incentivizing Reasoning and Self-Reflection Capacity for VLA Model in Autonomous Driving

DGX agent

arXiv:2509.01944v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models in autonomous driving systems have recently demonstrated transformative potential by integrating multimoda

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

Automating Crash Diagram Generation Using Vision-Language Models: A Case Study on Multi-Lane Roundabouts

DGX agent

arXiv:2604.15332v1 Announce Type: cross Abstract: Crash diagrams are essential tools in transportation safety analysis, yet their manual preparation remains time-consuming and prone to human variabili

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Concept-wise Attention for Fine-grained Concept Bottleneck Models

DGX agent

arXiv:2604.15748v1 Announce Type: new Abstract: Recently impressive performance has been achieved in Concept Bottleneck Models (CBM) by utilizing the image-text alignment learned by a large pre-traine

safetyarxiv-cs-cv
20 Apr 2026
Model Releases

ConFu: Contemplate the Future for Better Speculative Sampling

DGX agent

arXiv:2603.08899v2 Announce Type: replace Abstract: Speculative decoding has emerged as a powerful approach to accelerate large language model (LLM) inference by employing lightweight draft models to

model-releasesarxiv-cs-cl
20 Apr 2026
Local Ai

DINOv3 Beats Specialized Detectors: A Simple Foundation Model Baseline for Image Forensics

DGX agent

arXiv:2604.16083v1 Announce Type: new Abstract: With the rapid advancement of deep generative models, realistic fake images have become increasingly accessible, yet existing localization methods rely

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

FETAL-GAUGE: A Benchmark for Assessing Vision-Language Models in Fetal Ultrasound

DGX agent

arXiv:2512.22278v2 Announce Type: replace Abstract: The growing demand for prenatal ultrasound imaging has intensified a global shortage of trained sonographers, creating barriers to essential fetal h

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly g…

DGX agent

Free ~20-50% tok/s on a local llama.cpp setup if you already have a draft model sharing vocabulary with your main one. Local stack quietly got faster this weekend https://x.com/TechIno219886/status/20

model-releasesclem-delangue--x
20 Apr 2026
Model Releases

HyperGVL: Benchmarking and Improving Large Vision-Language Models in Hypergraph Understanding and Reasoning

DGX agent

arXiv:2604.15648v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) consistently require new arenas to guide their expanding boundaries, yet their capabilities with hypergraphs remain

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up t…

DGX agent

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up to 3x the tokens for images - it's priced the same as Opus 4.

model-releasesjeremy-howard--x
20 Apr 2026
Safety

Jailbreak Scaling Laws for Large Language Models: Polynomial-Exponential Crossover

DGX agent

arXiv:2603.11331v2 Announce Type: replace-cross Abstract: Adversarial attacks can reliably steer safety-aligned large language models toward unsafe behavior. Empirically, we find that strong adversari

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

JumpLoRA: Sparse Adapters for Continual Learning in Large Language Models

DGX agent

arXiv:2604.16171v1 Announce Type: cross Abstract: Adapter-based methods have become a cost-effective approach to continual learning (CL) for Large Language Models (LLMs), by sequentially learning a lo

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on mu…

DGX agent

life when you discover an open-source model that runs 300 parallel agents, executes for 12+ hours straight, beats GPT-5.4 and opus 4.6 on multiple benchmarks... and the weights are on huggingface Medi

model-releasesclem-delangue--x
20 Apr 2026
Model Releases

Reasoning-targeted Jailbreak Attacks on Large Reasoning Models via Semantic Triggers and Psychological Framing

DGX agent

arXiv:2604.15725v1 Announce Type: cross Abstract: Large Reasoning Models (LRMs) have demonstrated strong capabilities in generating step-by-step reasoning chains alongside final answers, enabling thei

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

RedBench: A Universal Dataset for Comprehensive Red Teaming of Large Language Models

DGX agent

arXiv:2601.03699v2 Announce Type: replace Abstract: As large language models (LLMs) become integral to safety-critical applications, ensuring their robustness against adversarial prompts is paramount.

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

Seed1.8 Model Card: Towards Generalized Real-World Agency

DGX agent

arXiv:2603.20633v3 Announce Type: replace Abstract: We present Seed1.8, a foundation model aimed at generalized real-world agency: going beyond single-turn prediction to multi-turn interaction, tool u

model-releasesarxiv-cs-ai
20 Apr 2026
Research

SSMamba: A Self-Supervised Hybrid State Space Model for Pathological Image Classification

DGX agent

arXiv:2604.15711v1 Announce Type: cross Abstract: Pathological diagnosis is highly reliant on image analysis, where Regions of Interest (ROIs) serve as the primary basis for diagnostic evidence, while

researcharxiv-cs-ai
20 Apr 2026
Model Releases

Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints

DGX agent

arXiv:2604.15664v1 Announce Type: new Abstract: The rise of autonomous AI agents suggests that dynamic benchmark environments with built-in feedback on scientifically grounded tasks are needed to eval

model-releasesarxiv-cs-lg
20 Apr 2026
Agents

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-t…

DGX agent

there’s clearly some confusing conflicts + double-think going on in AI between: 1. Closed labs saying use our harness, it’s naturally post-trained and gets the best out of models by being “in-distribu

agentsharrison-chase--x
20 Apr 2026
Safety

Towards Intrinsic Interpretability of Large Language Models:A Survey of Design Principles and Architectures

DGX agent

arXiv:2604.16042v1 Announce Type: cross Abstract: While Large Language Models (LLMs) have achieved strong performance across many NLP tasks, their opaque internal mechanisms hinder trustworthiness and

safetyarxiv-cs-ai
20 Apr 2026
Applications

Unveiling Stochasticity: Universal Multi-modal Probabilistic Modeling for Traffic Forecasting

DGX agent

arXiv:2604.16084v1 Announce Type: cross Abstract: Traffic forecasting is a challenging spatio-temporal modeling task and a critical component of urban transportation management. Current studies mainly

applicationsarxiv-cs-ai
20 Apr 2026
Model Releases

When Surfaces Lie: Exploiting Wrinkle-Induced Attention Shift to Attack Vision-Language Models

DGX agent

arXiv:2603.27759v3 Announce Type: replace Abstract: Visual-Language Models (VLMs) have demonstrated exceptional cross-modal understanding across various tasks, including zero-shot classification, imag

model-releasesarxiv-cs-cv
20 Apr 2026
Model Releases

An obvious way to release Mythos class models with uncertain autonomous ability is to make them only available on the website, like Gemini D…

DGX agent

An obvious way to release Mythos class models with uncertain autonomous ability is to make them only available on the website, like Gemini Deep Think or ChatGPT Pro. Minimal risk of being used for aut

model-releasesethan-mollick--x
19 Apr 2026
Model Releases

LiteParse is the best model-free, open-source document parser for AI agents. It now gets a first-class landing page on our website 💫 Our co…

DGX agent

LiteParse is the best model-free, open-source document parser for AI agents. It now gets a first-class landing page on our website 💫 Our company mission is building the world's best agentic document p

model-releasesjerry-liu--x
19 Apr 2026
Model Releases

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your ag…

DGX agent

Hermes Agent is model & tool backend agnostic for a reason, everyone should have access to AI. We don't dictate the rules of use for your agent, YOU do Anthropic shut down an entire company's Claude a

model-releasesnous-research--x
18 Apr 2026
Research

An Analysis of Regularization and Fokker-Planck Residuals in Diffusion Models for Image Generation

DGX agent

arXiv:2604.15171v1 Announce Type: new Abstract: Recent work has shown that diffusion models trained with the denoising score matching (DSM) objective often violate the Fokker--Planck (FP) equation tha

researcharxiv-cs-cv
17 Apr 2026
Agents

Empowerment Gain and Causal Model Construction: Children and adults are sensitive to controllability and variability in their causal interventions

DGX agent

arXiv:2512.08230v2 Announce Type: replace Abstract: Learning about the causal structure of the world is a fundamental problem for human cognition. Causal models and especially causal learning have pro

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

EuropeMedQA Study Protocol: A Multilingual, Multimodal Medical Examination Dataset for Language Model Evaluation

DGX agent

arXiv:2604.14306v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated high proficiency on English-centric medical examinations, their performance often declines when fac

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

GraphScout: Empowering Large Language Models with Intrinsic Exploration Ability for Agentic Graph Reasoning

DGX agent

arXiv:2603.01410v2 Announce Type: replace Abstract: Knowledge graphs provide structured and reliable information for many real-world applications, motivating increasing interest in combining large lan

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

Improving Language Models with Intentional Analysis

DGX agent

arXiv:2502.04689v4 Announce Type: replace Abstract: Intent, a critical cognitive notion and mental state, is ubiquitous in human communication and problem-solving. Accurately understanding the underly

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

It's never been easier for your agent to harness the power of Hugging Face, from account creation to getting your own finetuned model, with …

DGX agent

It's never been easier for your agent to harness the power of Hugging Face, from account creation to getting your own finetuned model, with everything in between, without any manual action on your end

agentsclem-delangue--x
17 Apr 2026
Safety

Language of Thought Shapes Output Diversity in Large Language Models

DGX agent

arXiv:2601.11227v2 Announce Type: replace Abstract: Output diversity is crucial for Large Language Models as it underpins pluralism and creativity. In this work, we reveal that controlling the languag

safetyarxiv-cs-cl
17 Apr 2026
Safety

Multi-Persona Thinking for Bias Mitigation in Large Language Models

DGX agent

arXiv:2601.15488v2 Announce Type: replace Abstract: Large Language Models (LLMs) exhibit social biases, which can lead to harmful stereotypes and unfair outcomes. We propose extbf{Multi-Persona Thinki

safetyarxiv-cs-cl
17 Apr 2026
Model Releases

Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an 'early sign' of generalization, surprising researchers (Connie Loizos/TechCrunch)

DGX agent

Connie Loizos / TechCrunch: Physical Intelligence says its new model, π0.7, can direct robots on tasks they weren't trained on, an “early sign” of generalization, surprising researchers — Physical Int

model-releasestechmeme
17 Apr 2026
Tutorials

Pruning Long Chain-of-Thought of Large Reasoning Models via Small-Scale Preference Optimization

DGX agent

arXiv:2508.10164v2 Announce Type: replace Abstract: Recent advances in Large Reasoning Models (LRMs) have demonstrated strong performance on complex tasks through long Chain-of-Thought (CoT) reasoning

tutorialsarxiv-cs-ai
17 Apr 2026
Research

Random Matrix Theory for Deep Learning: Beyond Eigenvalues of Linear Models

DGX agent

arXiv:2506.13139v2 Announce Type: replace-cross Abstract: Modern Machine Learning (ML) and Deep Neural Networks (DNNs) often operate on high-dimensional data and rely on overparameterized models, wher

researcharxiv-cs-lg
17 Apr 2026
Safety

RaTA-Tool: Retrieval-based Tool Selection with Multimodal Large Language Models

DGX agent

arXiv:2604.14951v1 Announce Type: cross Abstract: Tool learning with foundation models aims to endow AI systems with the ability to invoke external resources -- such as APIs, computational utilities,

safetyarxiv-cs-cl
17 Apr 2026
Safety

RECOVER: Designing a Large Language Model-based Remote Patient Monitoring System for Postoperative Gastrointestinal Cancer Care

DGX agent

arXiv:2502.05740v2 Announce Type: replace-cross Abstract: Cancer surgery is a key treatment for gastrointestinal (GI) cancers, a group of cancers that account for more than 35% of cancer-related death

safetyarxiv-cs-ai
17 Apr 2026
Applications

Route to Rome Attack: Directing LLM Routers to Expensive Models via Adversarial Suffix Optimization

DGX agent

arXiv:2604.15022v1 Announce Type: cross Abstract: Cost-aware routing dynamically dispatches user queries to models of varying capability to balance performance and inference cost. However, the routing

applicationsarxiv-cs-cl
17 Apr 2026
Model Releases

SelfGrader: Stable Jailbreak Detection for Large Language Models using Token-Level Logits

DGX agent

arXiv:2604.01473v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are powerful tools for answering user queries, yet they remain highly vulnerable to jailbreak attacks. Existing g

model-releasesarxiv-cs-ai
17 Apr 2026
← Previous
1…150151152153154…1262
Next →