AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
60,405 results
Safety

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

DGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Parallel Test-Time Scaling for Latent Reasoning Models

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2510.07745v4 Announce Type: replace Abstract: Parallel test-time scaling (TTS) is a pivotal approach for enhancing large language models (LLMs), typically by sampling multiple token-based chains

model-releasesarxiv-cs-cl
21 Apr 2026
Tutorials

Plausibility as Commonsense Reasoning: Humans Succeed, Large Language Models Do not

DGX agent

arXiv:2604.04825v2 Announce Type: replace Abstract: Large language models achieve strong performance on many language tasks, yet it remains unclear whether they integrate world knowledge with syntacti

tutorialsarxiv-cs-cl
21 Apr 2026
Safety

Privacy Collapse: Benign Fine-Tuning Can Break Contextual Privacy in Language Models

DGX agent

arXiv:2601.15220v2 Announce Type: replace Abstract: We identify a novel phenomenon in language models: benign fine-tuning of frontier models can lead to privacy collapse. We find that diverse, subtle

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

Program Structure-aware Language Models: Targeted Software Testing beyond Textual Semantics

DGX agent

arXiv:2604.17715v1 Announce Type: cross Abstract: Recent advances in large language models for test case generation have improved branch coverage via prompt-engineered mutations. However, they still l

model-releasesarxiv-cs-lg
21 Apr 2026
Research

Residual Diffusion Bridge Model for Image Restoration

DGX agent

arXiv:2510.23116v3 Announce Type: replace Abstract: Diffusion bridge models establish probabilistic paths between arbitrary paired distributions and exhibit great potential for universal image restora

researcharxiv-cs-cv
21 Apr 2026
Model Releases

Revisiting Change VQA in Remote Sensing with Structured and Native Multimodal Qwen Models

DGX agent

arXiv:2604.18429v1 Announce Type: new Abstract: Change visual question answering (Change VQA) addresses the problem of answering natural-language questions about semantic changes between bi-temporal r

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

S2H-DPO: Hardness-Aware Preference Optimization for Vision-Language Models

DGX agent

arXiv:2604.18512v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have demonstrated remarkable progress in single-image understanding, yet effective reasoning across multiple images remain

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Sky2Ground: A Benchmark for Site Modeling under Varying Altitude

DGX agent

arXiv:2603.13740v3 Announce Type: replace Abstract: We introduce Sky2Ground, a three-view dataset designed for varying altitude camera localization, correspondence learning, and reconstruction. The da

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Test-Time Adaptation for EEG Foundation Models: A Systematic Study under Real-World Distribution Shifts

DGX agent

arXiv:2604.16926v1 Announce Type: new Abstract: Electroencephalography (EEG) foundation models have shown strong potential for learning generalizable representations from large-scale neural data, yet

model-releasesarxiv-cs-lg
21 Apr 2026
Research

When Background Matters: Breaking Medical Vision Language Models by Transferable Attack

DGX agent

arXiv:2604.17318v1 Announce Type: new Abstract: Vision-Language Models (VLMs) are increasingly used in clinical diagnostics, yet their robustness to adversarial attacks remains largely unexplored, pos

researcharxiv-cs-cv
21 Apr 2026
Safety

When Choices Become Risks: Safety Failures of Large Language Models under Multiple-Choice Constraints

DGX agent

arXiv:2604.16916v1 Announce Type: new Abstract: Safety alignment in large language models (LLMs) is primarily evaluated under open-ended generation, where models can mitigate risk by refusing to respo

safetyarxiv-cs-cl
21 Apr 2026
Research

When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models

DGX agent

arXiv:2507.13868v2 Announce Type: replace Abstract: Vision-language models (VLMs) increasingly combine visual and textual information to perform complex tasks. However, conflicts between their interna

researcharxiv-cs-cv
21 Apr 2026
Local Ai

AdaVFM: Adaptive Vision Foundation Models for Edge Intelligence via LLM-Guided Execution

DGX agent

arXiv:2604.15622v1 Announce Type: new Abstract: Language-aligned vision foundation models (VFMs) enable versatile visual understanding for always-on contextual AI, but their deployment on edge devices

local-aiarxiv-cs-cv
20 Apr 2026
Model Releases

ChatENV: An Interactive Vision-Language Model for Sensor-Guided Environmental Monitoring and Scenario Simulation

DGX agent

arXiv:2508.10635v3 Announce Type: replace Abstract: Understanding environmental changes from remote sensing imagery is vital for climate resilience, urban planning, and ecosystem monitoring. Yet, curr

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

FineSteer: A Unified Framework for Fine-Grained Inference-Time Steering in Large Language Models

DGX agent

arXiv:2604.15488v1 Announce Type: cross Abstract: Large language models (LLMs) often exhibit undesirable behaviors, such as safety violations and hallucinations. Although inference-time steering offer

safetyarxiv-cs-ai
20 Apr 2026
Research

How Hypocritical Is Your LLM judge? Listener-Speaker Asymmetries in the Pragmatic Competence of Large Language Models

DGX agent

arXiv:2604.15873v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly studied as repositories of linguistic knowledge. In this line of work, models are commonly evaluated both

researcharxiv-cs-cl
20 Apr 2026
Safety

Improving Reasoning Capabilities in Small Models through Mixture-of-Layers Distillation with Stepwise Attention on Key Information

DGX agent

arXiv:2604.15701v1 Announce Type: new Abstract: The significant computational demands of large language models have increased interest in distilling reasoning abilities into smaller models via Chain-o

safetyarxiv-cs-cl
20 Apr 2026
Model Releases

Preference Estimation via Opponent Modeling in Multi-Agent Negotiation

DGX agent

arXiv:2604.15687v1 Announce Type: new Abstract: Automated negotiation in complex, multi-party and multi-issue settings critically depends on accurate opponent modeling. However, conventional numerical

model-releasesarxiv-cs-cl
20 Apr 2026
Model Releases

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This Mo…

DGX agent

🤯QWEN 3.6 35B-A3B IS INSANE AT CODING @unslothai recently dropped Qwen3.6-35B-A3B-GGUF The first open Qwen3.6 model and it’s solid. This MoE beast is cooking on benchmarks: 🧠SWE-bench Verified: 73.4 (

model-releasesclem-delangue--x
20 Apr 2026
Model Releases

Repurposing 3D Generative Model for Autoregressive Layout Generation

DGX agent

arXiv:2604.16299v1 Announce Type: new Abstract: We introduce LaviGen, a framework that repurposes 3D generative models for 3D layout generation. Unlike previous methods that infer object layouts from

model-releasesarxiv-cs-cv
20 Apr 2026
Applications

Tabular foundation models for in-context prediction of molecular properties

DGX agent

arXiv:2604.16123v1 Announce Type: new Abstract: Accurate molecular property prediction is central to drug discovery, catalysis, and process design, yet real-world applications are often limited by sma

applicationsarxiv-cs-lg
20 Apr 2026
Model Releases

I gave two MoE models the same vibe coding challenge Qwen3.6 35B A3B (31.8GB) vs Gemma4 26B A4B (23.3GB) Stack: > Unsloth Q6_K_XL > llama.cp…

DGX agent

I gave two MoE models the same vibe coding challenge Qwen3.6 35B A3B (31.8GB) vs Gemma4 26B A4B (23.3GB) Stack: > Unsloth Q6_K_XL > llama.cpp > Model-card recommended sampling for each 4 prompts, side

model-releasesclem-delangue--x
19 Apr 2026
Tutorials

ICYMI from a few weeks back, we compiled our learnings around how to achieve Training-Inference Parity in MoE Models. The Fundamental Issue:…

DGX agent

Fireworks AI published a compilation of insights on achieving training-inference parity in Mixture of Experts (MoE) models, addressing fundamental challenges in aligning model behavior between trainin

tutorialsfireworks-ai--x
18 Apr 2026
Model Releases

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite d…

DGX agent

Okay this one is insane. A new 18B frankenstein model was just released on @huggingface — Beats the new Qwen3.6-35B-A3B on a 44-test suite despite requiring 12GB VRAM instead of 24GB 🤯 Runs on a SINGL

model-releasesclem-delangue--x
18 Apr 2026
Model Releases

Anthropic’s new cybersecurity model could get it back in the government’s good graces

DGX agent

The Trump administration has spent nearly two months fighting with AI company Anthropic. It's dubbed the company a 'RADICAL LEFT, WOKE COMPANY' full of 'Leftwing nut jobs' and a menace to national sec

model-releasesthe-verge-ai
17 Apr 2026
Research

Assessing the Performance-Efficiency Trade-off of Foundation Models in Probabilistic Electricity Price Forecasting

DGX agent

arXiv:2604.14739v1 Announce Type: new Abstract: Large-scale renewable energy deployment introduces pronounced volatility into the electricity system, turning grid operation into a complex stochastic o

researcharxiv-cs-lg
17 Apr 2026
Local Ai

Best Ollama model for n8n workflows (RAG, file handling, reasoning) + hardware requirements?

DGX agent

Models like Qwen3 and Llama 3.2 are commonly used for n8n RAG workflows , with selection depending on use case requirements. Running Ollama models locally requires at least 16 GB of RAM on your device

local-air-ollama
17 Apr 2026
Model Releases

CausalDetox: Causal Head Selection and Intervention for Language Model Detoxification

DGX agent

arXiv:2604.14602v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate toxic content, posing significant risks for safe deployment. Current mitigation strategies often degrad

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

CobwebTM: Probabilistic Concept Formation for Lifelong and Hierarchical Topic Modeling

DGX agent

arXiv:2604.14489v1 Announce Type: new Abstract: Topic modeling seeks to uncover latent semantic structure in text corpora with minimal supervision. Neural approaches achieve strong performance but req

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’r…

DGX agent

Data Driven Agent Design with Evals & Hill Climbing Algorithms this is a mental model dump i’ve been thinking through + iterating on as we’re building self-improvement infra around agents: - mining Tr

agentsharrison-chase--x
17 Apr 2026
Model Releases

DharmaOCR: Specialized Small Language Models for Structured OCR that outperform Open-Source and Commercial Baselines

DGX agent

arXiv:2604.14314v1 Announce Type: cross Abstract: This manuscript introduces DharmaOCR Full and Lite, a pair of specialized small language models (SSLMs) for structured OCR that jointly optimize trans

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Domain-Adaptive Model Merging Across Disconnected Modes

DGX agent

arXiv:2603.05957v2 Announce Type: replace-cross Abstract: Learning across domains is challenging when data cannot be centralized due to privacy or heterogeneity, which limits the ability to train a si

researcharxiv-cs-ai
17 Apr 2026
Research

Exploring the flavor structure of leptons via diffusion models

DGX agent

arXiv:2503.21432v2 Announce Type: replace-cross Abstract: We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelli

researcharxiv-cs-lg
17 Apr 2026
Safety

Generative Models and Connected and Automated Vehicles: A Survey in Exploring the Intersection of Transportation and AI

DGX agent

arXiv:2403.10559v3 Announce Type: replace Abstract: This report investigates the history and impact of Generative Models and Connected and Automated Vehicles (CAVs), two groundbreaking forces pushing

safetyarxiv-cs-lg
17 Apr 2026
Research

HY-World 2.0: A Multi-Modal World Model for Reconstructing, Generating, and Simulating 3D Worlds

DGX agent

arXiv:2604.14268v1 Announce Type: new Abstract: We introduce HY-World 2.0, a multi-modal world model framework that advances our prior project HY-World 1.0. HY-World 2.0 accommodates diverse input mod

researcharxiv-cs-cv
17 Apr 2026
Model Releases

Prompt-Guided Image Editing with Masked Logit Nudging in Visual Autoregressive Models

DGX agent

arXiv:2604.14591v1 Announce Type: new Abstract: We address the problem of prompt-guided image editing in visual autoregressive models. Given a source image and a target text prompt, we aim to modify t

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

ProRank: Prompt Warmup via Reinforcement Learning for Small Language Models Reranking

DGX agent

arXiv:2506.03487v3 Announce Type: replace-cross Abstract: Reranking is fundamental to information retrieval and retrieval-augmented generation, with recent Large Language Models (LLMs) significantly a

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Psychological Steering of Large Language Models

DGX agent

arXiv:2604.14463v1 Announce Type: new Abstract: Large language models (LLMs) emulate a consistent human-like behavior that can be shaped through activation-level interventions. This paradigm is conver

researcharxiv-cs-cl
17 Apr 2026
Model Releases

QuantCode-Bench: A Benchmark for Evaluating the Ability of Large Language Models to Generate Executable Algorithmic Trading Strategies

DGX agent

arXiv:2604.15151v1 Announce Type: new Abstract: Large language models have demonstrated strong performance on general-purpose programming tasks, yet their ability to generate executable algorithmic tr

model-releasesarxiv-cs-cl
17 Apr 2026
Research

Style Amnesia: Investigating Speaking Style Degradation and Mitigation in Multi-Turn Spoken Language Models

DGX agent

arXiv:2512.23578v3 Announce Type: replace Abstract: In this paper, we show that when spoken language models (SLMs) are instructed to speak in a specific speaking style at the beginning of a multi-turn

researcharxiv-cs-cl
17 Apr 2026
Model Releases

Claude Opus 4.7 is now the default orchestration model powering Computer. It's also available for Max subscribers on Perplexity web, iOS, an…

DGX agent

Claude Opus 4.7 has been set as the default orchestration model for Anthropic's Computer product. The model is also available to Max subscribers on Perplexity's web platform and iOS application.

model-releasesperplexity--x
16 Apr 2026
Model Releases

fMRI-LM: Towards a Universal Foundation Model for Language-Aligned fMRI Understanding

DGX agent

arXiv:2511.21760v3 Announce Type: replace Abstract: Recent advances in multimodal large language models (LLMs) have enabled unified reasoning across images, audio, and video, but extending such capabi

model-releasesarxiv-cs-cl
16 Apr 2026
Model Releases

Google’s DeepMind just released new 4B and 27B MedGemma models!

DGX agent

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

model-releasesr-ollama
16 Apr 2026
Model Releases

L2D-Clinical: Learning to Defer for Adaptive Model Selection in Clinical Text Classification

DGX agent

arXiv:2604.13285v1 Announce Type: new Abstract: Clinical text classification requires choosing between specialized fine-tuned models (BERT variants) and general-purpose large language models (LLMs), y

model-releasesarxiv-cs-cl
16 Apr 2026
Industry

Model 3 & Y are the most cost-effective vehicles you can own

DGX agent

Model 3 & Y are the most cost-effective vehicles you can own This is a great graphic that compares cost per kilometer in Australia between the top 10 ICE vehicles & top 10 EVs. The difference is huge.

industryelon-musk--x
16 Apr 2026
Research

ReproMIA: A Comprehensive Analysis of Model Reprogramming for Proactive Membership Inference Attacks

DGX agent

arXiv:2603.28942v3 Announce Type: replace Abstract: The pervasive deployment of deep learning models across critical domains has concurrently intensified privacy concerns due to their inherent propens

researcharxiv-cs-lg
16 Apr 2026
Model Releases

Reward Hacking in the Era of Large Models: Mechanisms, Emergent Misalignment, Challenges

DGX agent

arXiv:2604.13602v1 Announce Type: new Abstract: Reinforcement Learning from Human Feedback (RLHF) and related alignment paradigms have become central to steering large language models (LLMs) and multi

model-releasesarxiv-cs-lg
16 Apr 2026
← Previous
1…7172737475…1259
Next →