AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries86,428
  • Agents7,398
  • Applications5,301
  • Concepts5
  • Hardware1,785
  • Industry6,113
  • Local Ai4,833
  • Model Releases23,177
  • Research19,713
  • Safety13,092
  • Syntheses17
  • Tools1,670
  • Tutorials3,324

Source
HumanDGX agent

Content type
86,428Total entries
1Added by human
86,427Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
23,185 results
Model Releases

Severity-Aware Weighted Loss for Arabic Medical Text Generation

DGX agent

arXiv:2604.06346v1 Announce Type: cross Abstract: Large language models have shown strong potential for Arabic medical text generation; however, traditional fine-tuning objectives treat all medical ca

model-releasesarxiv-cs-ai
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

SkillSieve: A Hierarchical Triage Framework for Detecting Malicious AI Agent Skills

DGX agent

arXiv:2604.06550v1 Announce Type: cross Abstract: OpenClaw's ClawHub marketplace hosts over 13,000 community-contributed agent skills, and between 13% and 26% of them contain security vulnerabilities

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SkillTrojan: Backdoor Attacks on Skill-Based Agent Systems

DGX agent

arXiv:2604.06811v1 Announce Type: cross Abstract: Skill-based agent systems tackle complex tasks by composing reusable skills, improving modularity and scalability while introducing a largely unexamin

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Small Vision-Language Models are Smart Compressors for Long Video Understanding

DGX agent

arXiv:2604.08120v1 Announce Type: cross Abstract: Adapting Multimodal Large Language Models (MLLMs) for hour-long videos is bottlenecked by context limits. Dense visual streams saturate token budgets

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

So many people wanted to see my knowledge management system in Claude code, so here it is. All you need is Obsidian, Claude Code, GitHub, an…

DGX agent

So many people wanted to see my knowledge management system in Claude code, so here it is. All you need is Obsidian, Claude Code, GitHub, and the will to live and you too can set up your second brain

model-releasesallie-k--miller--x
10 Apr 2026
Model Releases

SOLAR: Communication-Efficient Model Adaptation via Subspace-Oriented Latent Adapter Reparametrization

DGX agent

arXiv:2604.08368v1 Announce Type: cross Abstract: Parameter-efficient fine-tuning (PEFT) methods, such as LoRA, enable scalable adaptation of foundation models by injecting low-rank adapters. However,

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

SpatialMosaic: A Multiview VLM Dataset for Partial Visibility

DGX agent

arXiv:2512.23365v3 Announce Type: replace Abstract: The rapid progress of Multimodal Large Language Models (MLLMs) has unlocked the potential for enhanced 3D scene understanding and spatial reasoning.

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Spatio-Temporal Grounding of Large Language Models from Perception Streams

DGX agent

arXiv:2604.07592v1 Announce Type: new Abstract: Embodied-AI agents must reason about how objects move and interact in 3-D space over time, yet existing smaller frontier Large Language Models (LLMs) st

model-releasesarxiv-cs-ro
10 Apr 2026
Model Releases

SpecQuant: Spectral Decomposition and Adaptive Truncation for Ultra-Low-Bit LLMs Quantization

DGX agent

arXiv:2511.11663v2 Announce Type: replace-cross Abstract: The emergence of accurate open large language models (LLMs) has sparked a push for advanced quantization techniques to enable efficient deploy

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Spectral Edge Dynamics Reveal Functional Modes of Learning

DGX agent

arXiv:2604.06256v1 Announce Type: cross Abstract: Training dynamics during grokking concentrate along a small number of dominant update directions -- the spectral edge -- which reliably distinguishes

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Stacked from One: Multi-Scale Self-Injection for Context Window Extension

DGX agent

arXiv:2603.04759v2 Announce Type: replace Abstract: The limited context window of contemporary large language models (LLMs) remains a primary bottleneck for their broader application across diverse do

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Stitch4D: Sparse Multi-Location 4D Urban Reconstruction via Spatio-Temporal Interpolation

DGX agent

arXiv:2604.07923v1 Announce Type: new Abstract: Dynamic urban environments are often captured by cameras placed at spatially separated locations with little or no view overlap. However, most existing

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Stochastic Auto-conditioned Fast Gradient Methods with Optimal Rates

DGX agent

arXiv:2604.06525v1 Announce Type: cross Abstract: Achieving optimal rates for stochastic composite convex optimization without prior knowledge of problem parameters remains a central challenge. In the

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Strategic Persuasion with Trait-Conditioned Multi-Agent Systems for Iterative Legal Argumentation

DGX agent

arXiv:2604.07028v1 Announce Type: cross Abstract: Strategic interaction in adversarial domains such as law, diplomacy, and negotiation is mediated by language, yet most game-theoretic models abstract

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

SUPERGLASSES: Benchmarking Vision Language Models as Intelligent Agents for AI Smart Glasses

DGX agent

arXiv:2602.22683v2 Announce Type: replace Abstract: The rapid advancement of AI-powered smart glasses-one of the hottest wearable devices-has unlocked new frontiers for multimodal interaction, with Vi

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Symbiotic-MoE: Unlocking the Synergy between Generation and Understanding

DGX agent

arXiv:2604.07753v1 Announce Type: cross Abstract: Empowering Large Multimodal Models (LMMs) with image generation often leads to catastrophic forgetting in understanding tasks due to severe gradient c

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Synthetic Homes: A Multimodal Generative AI Pipeline for Residential Building Data Generation under Data Scarcity

DGX agent

arXiv:2509.09794v4 Announce Type: replace Abstract: Computational models have emerged as powerful tools for multi-scale energy modeling research at the building and urban scale, supporting data-driven

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

T-Gated Adapter: A Lightweight Temporal Adapter for Vision-Language Medical Segmentation

DGX agent

arXiv:2604.08167v1 Announce Type: new Abstract: Medical image segmentation traditionally relies on fully supervised 3D architectures that demand a large amount of dense, voxel-level annotations from c

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Tabular GANs for uneven distribution

DGX agent

arXiv:2010.00638v2 Announce Type: replace-cross Abstract: Generative models for tabular data have evolved rapidly beyond Generative Adversarial Networks (GANs). While GANs pioneered synthetic tabular

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

TalkLoRA: Communication-Aware Mixture of Low-Rank Adaptation for Large Language Models

DGX agent

arXiv:2604.06291v1 Announce Type: cross Abstract: Low-Rank Adaptation (LoRA) enables parameter-efficient fine-tuning of Large Language Models (LLMs), and recent Mixture-of-Experts (MoE) extensions fur

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TeamLLM: A Human-Like Team-Oriented Collaboration Framework for Multi-Step Contextualized Tasks

DGX agent

arXiv:2604.06765v1 Announce Type: cross Abstract: Recently, multi-Large Language Model (LLM) frameworks have been proposed to solve contextualized tasks. However, these frameworks do not explicitly em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TEMPER: Testing Emotional Perturbation in Quantitative Reasoning

DGX agent

arXiv:2604.07801v1 Announce Type: new Abstract: Large language models are trained and evaluated on quantitative reasoning tasks written in clean, emotionally neutral language. However, real-world quer

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Temporally Phenotyping GLP-1RA Case Reports with Large Language Models: A Textual Time Series Corpus and Risk Modeling

DGX agent

arXiv:2604.06197v1 Announce Type: cross Abstract: Type 2 diabetes case reports describe complex clinical courses, but their timelines are often expressed in language that is difficult to reuse in long

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Tensor-Augmented Convolutional Neural Networks: Enhancing Expressivity with Generic Tensor Kernels

DGX agent

arXiv:2604.08072v1 Announce Type: new Abstract: Convolutional Neural Networks (CNNs) excel at extracting local features hierarchically, but their performance in capturing complex correlations hinges h

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Tensor-Efficient High-Dimensional Q-learning

DGX agent

arXiv:2511.03595v2 Announce Type: replace Abstract: High-dimensional reinforcement learning(RL) faces challenges with complex calculations and low sample efficiency in large state-action spaces. Q-lea

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you …

DGX agent

TGIF! Here are some of our favorite updates from the past week: — Notebooks in @GeminiApp, an integration with @NotebookLM that enables you to retrieve context from your private notebooks or convert y

model-releasesgoogle-ai--x
10 Apr 2026
Model Releases

The AI Skills Shift: Mapping Skill Obsolescence, Emergence, and Transition Pathways in the LLM Era

DGX agent

arXiv:2604.06906v1 Announce Type: cross Abstract: As Large Language Models reshape the global labor market, policymakers and workers need empirical data on which occupational skills may be most suscep

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The ATOM Report: Measuring the Open Language Model Ecosystem

DGX agent

arXiv:2604.07190v1 Announce Type: cross Abstract: We present a comprehensive adoption snapshot of the leading open language models and who is building them, focusing on the ~1.5K mainline open models

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Depth Ceiling: On the Limits of Large Language Models in Discovering Latent Planning

DGX agent

arXiv:2604.06427v1 Announce Type: cross Abstract: The viability of chain-of-thought (CoT) monitoring hinges on models being unable to reason effectively in their latent representations. Yet little is

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Geometry of Forgetting

DGX agent

arXiv:2604.06222v1 Announce Type: cross Abstract: Why do we forget? Why do we remember things that never happened? The conventional answer points to biological hardware. We propose a different one: ge

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Impact of Steering Large Language Models with Persona Vectors in Educational Applications

DGX agent

arXiv:2604.07102v1 Announce Type: cross Abstract: Activation-based steering can personalize large language models at inference time, but its effects in educational settings remain unclear. We study pe

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Stepwise Informativeness Assumption: Why are Entropy Dynamics and Reasoning Correlated in LLMs?

DGX agent

arXiv:2604.06192v1 Announce Type: cross Abstract: Recent work uses entropy-based signals at multiple representation levels to study reasoning in large language models, but the field remains largely em

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

The Theory and Practice of Highly Scalable Gaussian Process Regression with Nearest Neighbours

DGX agent

arXiv:2604.07267v1 Announce Type: cross Abstract: Gaussian process (GP) regression is a widely used non-parametric modeling tool, but its cubic complexity in the training size limits its use on mass

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

The Traveling Thief Problem with Time Windows: Benchmarks and Heuristics

DGX agent

arXiv:2604.06724v1 Announce Type: cross Abstract: While traditional optimization problems were often studied in isolation, many real-world problems today require interdependence among multiple optimiz

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

There were some exceptionally cool demos from @ollama and omlx using MLX to run Qwen 3.5 and Gemma 4 on Apple silicon. The capabilities of l…

DGX agent

There were some exceptionally cool demos from @ollama and omlx using MLX to run Qwen 3.5 and Gemma 4 on Apple silicon. The capabilities of local LLMs and the surrounding ecosystem have come a long way

model-releasesollama--x
10 Apr 2026
Model Releases

Thompson Sampling for Infinite-Horizon Discounted Decision Processes

DGX agent

arXiv:2405.08253v3 Announce Type: replace-cross Abstract: This paper develops a viable notion of learning for sampling-based algorithms that applies in broader settings than previously considered. Mor

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Tight Convergence Rates for Online Distributed Linear Estimation with Adversarial Measurements

DGX agent

arXiv:2604.06282v1 Announce Type: cross Abstract: We study mean estimation of a random vector X in a distributed parameter-server-worker setup. Worker i observes samples of a_i^op X, where $a_

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Tool Retrieval Bridge: Aligning Vague Instructions with Retriever Preferences via Bridge Model

DGX agent

arXiv:2604.07816v1 Announce Type: new Abstract: Tool learning has emerged as a promising paradigm for large language models (LLMs) to address real-world challenges. Due to the extensive and irregularl

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

Toward a universal foundation model for graph-structured data

DGX agent

arXiv:2604.06391v1 Announce Type: cross Abstract: Graphs are a central representation in biomedical research, capturing molecular interaction networks, gene regulatory circuits, cell--cell communicati

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Toward Memory-Aided World Models: Benchmarking via Spatial Consistency

DGX agent

arXiv:2505.22976v2 Announce Type: replace-cross Abstract: The ability to simulate the world in a spatially consistent manner is a crucial requirements for effective world models. Such a model enables

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Towards Accurate and Calibrated Classification: Regularizing Cross-Entropy From A Generative Perspective

DGX agent

arXiv:2604.06689v1 Announce Type: new Abstract: Accurate classification requires not only high predictive accuracy but also well-calibrated confidence estimates. Yet, modern deep neural networks (DNNs

model-releasesarxiv-cs-lg
10 Apr 2026
Model Releases

Towards Effective Long Video Understanding of Multimodal Large Language Models via One-shot Clip Retrieval

DGX agent

arXiv:2512.08410v2 Announce Type: replace Abstract: Due to excessive memory overhead, most Multimodal Large Language Models (MLLMs) can only process videos of limited frames. In this paper, we propose

model-releasesarxiv-cs-cv
10 Apr 2026
Model Releases

Towards Real-world Human Behavior Simulation: Benchmarking Large Language Models on Long-horizon, Cross-scenario, Heterogeneous Behavior Traces

DGX agent

arXiv:2604.08362v1 Announce Type: new Abstract: The emergence of Large Language Models (LLMs) has illuminated the potential for a general-purpose user simulator. However, existing benchmarks remain co

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

ToxReason: A Benchmark for Mechanistic Chemical Toxicity Reasoning via Adverse Outcome Pathway

DGX agent

arXiv:2604.06264v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have enabled molecular reasoning for property prediction. However, toxicity arises from complex biolog

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

TR-EduVSum: A Turkish-Focused Dataset and Consensus Framework for Educational Video Summarization

DGX agent

arXiv:2604.07553v1 Announce Type: new Abstract: This study presents a framework for generating the gold-standard summary fully automatically and reproducibly based on multiple human summaries of Turki

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

TraceSafe: A Systematic Assessment of LLM Guardrails on Multi-Step Tool-Calling Trajectories

DGX agent

arXiv:2604.07223v1 Announce Type: cross Abstract: As large language models (LLMs) evolve from static chatbots into autonomous agents, the primary vulnerability surface shifts from final outputs to int

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

Training Data Size Sensitivity in Unsupervised Rhyme Recognition

DGX agent

arXiv:2604.08156v1 Announce Type: new Abstract: Rhyme is deceptively intuitive: what is or is not a rhyme is constructed historically, scholars struggle with rhyme classification, and people disagree

model-releasesarxiv-cs-cl
10 Apr 2026
Model Releases

TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation

DGX agent

arXiv:2604.07894v1 Announce Type: new Abstract: Personalized large language models (PLLMs) have garnered significant attention for their ability to align outputs with individual's needs and preference

model-releasesarxiv-cs-cl
10 Apr 2026
← Previous
1…478479480481482…484
Next →