AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
17,131 results
Model Releases

Does Forgetting Transfer Across Modalities? A Real-World Benchmark for Cross-Modal Knowledge Unlearning Evaluation

DGX agent

arXiv:2608.03791v1 Announce Type: new Abstract: Vision-Language Models (VLMs), like Large Language Models (LLMs), may memorize sensitive, copyrighted, or harmful knowledge from their pretraining corpo

model-releasesarxiv-cs-ai
5 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Don't Walk the Line: Boundary Guidance for Filtered Generation

DGX agent

arXiv:2510.11834v3 Announce Type: replace-cross Abstract: Generative models are increasingly paired with safety classifiers that filter harmful or undesirable outputs. A common strategy is to fine-tun

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

dots.tts.edit: Precisely Controlled Speech Editing with a Continuous Autoregressive Model

DGX agent

arXiv:2608.02673v1 Announce Type: cross Abstract: Speech editing for content creation requires precise control over both what an edit should do and where it should apply. Free-form natural language pr

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Double Descent in Gradient Boosting Decision Trees via Split-Candidate Scaling

DGX agent

arXiv:2608.03111v1 Announce Type: new Abstract: Double descent is commonly studied by scaling an explicit capacity parameter, such as neural-network width. For gradient boosting decision trees (GBDTs)

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

DP-MemView: A Memory Interface for Attribute-Level Transcript Privacy in Long-Term LLM Agents

DGX agent

arXiv:2608.03130v1 Announce Type: cross Abstract: Long-term memory enables persistent personalization in LLM agents, but repeated memory-conditioned responses can cumulatively reveal protected attribu

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

DS@GT-ARC at eRisk 2026 Task 3: Sparse, Semantic, and LLM Reranking for ADHD Symptom Sentences

DGX agent

arXiv:2608.03883v1 Announce Type: new Abstract: This paper describes our submissions to eRisk 2026 Task 3, ADHD Symptom Sentence Ranking. The task requires systems to rank candidate Reddit sentences a

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Dynamically Allocating Evaluation Effort for Model Ranking

DGX agent

arXiv:2608.03437v1 Announce Type: new Abstract: While human evaluation is the gold standard in many NLP tasks, it suffers from prohibitive costs and poor scalability. When identifying top-performing m

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

EditFlow3D: Automated Local Editing of 3D Assets with Trajectory Preservation

DGX agent

arXiv:2608.03179v1 Announce Type: new Abstract: Controllable local editing of 3D assets requires precise target localization and appropriate visual guidance. However, existing methods lack a simple ye

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

DGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Efficient Knowledge Distillation for LLMs: Offline Top-K Logits and a Fused Chunked KL Loss

DGX agent

arXiv:2608.03796v1 Announce Type: cross Abstract: Small language models are often the only option for deployment under tight latency, cost, and on-premises constraints, but they are rarely trained fro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Efficient Multilingual Neural Machine Translation via Corpus-Driven Vocabulary Pruning: An English-Arabic Case Study

DGX agent

arXiv:2608.03480v1 Announce Type: new Abstract: The adoption of large pre-trained multilingual models for neural machine translation (MNMT) faces a major challenge: excessive memory and computational

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Efficient unsupervised domain adaptation via self-supervised vision transformer and synergistic cross-domain alignment

DGX agent

arXiv:2407.21311v2 Announce Type: replace-cross Abstract: Unsupervised domain adaptation (UDA) aims to mitigate domain shift, where the distribution of labeled source data differs from that of unlabel

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Enhancing Tabular Learners with Context-Aware Semantic Embeddings

DGX agent

arXiv:2608.03565v1 Announce Type: new Abstract: While modern tabular learners excel at capturing statistical patterns, they frequently operate in a semantic vacuum, treating textual features as discre

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluating Counterfactual Sensitivity to Patient Information in Medication-Safety Reasoning

DGX agent

arXiv:2608.03028v1 Announce Type: new Abstract: Applying a valid medication-safety rule when its patient-specific conditions are not met can produce an incorrect decision. Existing medical evaluations

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluating LLM Trade-offs for Enterprise Automation: Lessons from Workflow Generation in a Production Enterprise Platform

DGX agent

arXiv:2608.03311v1 Announce Type: cross Abstract: Enterprise compliance management requires rapid adaptation to evolving regulatory frameworks (e.g., DORA, AI RMF, FedRAMP) and tight remediation SLAs.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluating LLMs in Database Scenarios: A Lifecycle Benchmark for Assessing Their Potential in Core Database Tasks

DGX agent

arXiv:2608.03794v1 Announce Type: cross Abstract: Large Language Models (LLMs) are transforming database interaction paradigms, evolving from simple query translators to autonomous database administra

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluating OpenAI's Privacy Filter: Cross-Lingual, Cross-Domain PII Detection Across 42 Benchmarks

DGX agent

arXiv:2608.02616v1 Announce Type: cross Abstract: We present the first independent, systematic evaluation of OpenAI's Privacy Filter (OPF), a 1.5B-parameter bidirectional PII detector, across 42 synth

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Evaluation Blindness: How Silent Measurement Failures Corrupt AI Systems from Training to Deployment

DGX agent

arXiv:2608.02786v1 Announce Type: new Abstract: AI systems can fail silently. The failure propagates through training loops, evaluation pipelines, and production monitoring stacks until downstream har

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Exploiting Separability in Multi-Scale Grey-Box Bayesian Optimization

DGX agent

arXiv:2608.03045v1 Announce Type: new Abstract: We consider grey-box optimization problems where the decision variables naturally partition into black-box variables (as arguments to an expensive black

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

Externally Validated Breast Ultrasound Segmentation via Multi-task Learning with BI-RADS-Consistent Morphological Priors

DGX agent

arXiv:2511.15968v2 Announce Type: replace-cross Abstract: External validation of breast ultrasound segmentation models remains limited because internal train--test splits do not capture domain shifts

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

DGX agent

arXiv:2608.03222v1 Announce Type: cross Abstract: Software engineering (SWE) agents resolve repository-level issues through long trajectories that grow increasingly expensive as context accumulates. F

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FakeI2V-Bench: Benchmarking the Applicability of Image-level Deepfake Detectors for Deepfake Video Detection

DGX agent

arXiv:2608.03096v1 Announce Type: cross Abstract: Recent advances in video generation models have significantly intensified the deepfake threat, yet the current deepfake video detection benchmarks rem

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FedCritic-MIMO: Communication-Efficient Serverless Federated Critic Learning for Massive-MIMO Resource Control in Open and Disaggregated 6G RANs

DGX agent

arXiv:2608.03852v1 Announce Type: new Abstract: This paper proposes FedCritic-MIMO, a communication-efficient serverless federated multi-agent reinforcement learning framework for AI-native resource c

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

FinVerse: Financial Time-Series Benchmark

DGX agent

arXiv:2608.03259v1 Announce Type: cross Abstract: As time-series foundation models have emerged, the need for benchmarks that can evaluate their forecasting ability in meaningful ways has become incre

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

FLARE: Few-shot Learning-based Adaptive Reflective Engine

DGX agent

arXiv:2608.02919v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in complex, compound AI systems where performance hinges on the quality of prompts. Recent state-

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

Forecasting Revenue with its Customer-Base Drivers: When and Why Coordination Helps

DGX agent

arXiv:2608.02911v1 Announce Type: new Abstract: Revenue forecasts guide acquisition budgets, demand planning, and customer-based valuations, yet an aggregate forecast does not show whether change refl

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

FreqAdapt: Frequency-Adaptive Processing for RAW Object Detection

DGX agent

arXiv:2608.03385v1 Announce Type: new Abstract: Existing object detection methods predominantly utilize sRGB inputs, which are compressed from RAW sensor data using Image Signal Processors (ISP) origi

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

From Generator to Embedder: Harnessing Innate Abilities of Multimodal LLMs via Building Zero-Shot Discriminative Embedding Model

DGX agent

arXiv:2508.00955v3 Announce Type: replace-cross Abstract: Adapting generative Multimodal Large Language Models (MLLMs) into universal embedding models typically demands resource-intensive contrastive

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

From Social Coding to Agentic Coding: Productivity and Relational Reconfiguration in Open-Source Communities

DGX agent

arXiv:2608.03585v1 Announce Type: new Abstract: Open-source software communities are a form of digital public infrastructure that not only produces code, but also generates public knowledge and interp

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Frozen High-Resolution Inference for Cross-City Object Detection: An AI City Challenge 2026 Study

DGX agent

arXiv:2608.03136v1 Announce Type: new Abstract: Cross-city object detection requires a detector trained in one city to generalize to an unlabeled target city. In AI City Challenge 2026 Track 6, we ana

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Fusion-Poly: A Polyhedral Framework Based on Spatial-Temporal Fusion for 3D Multi-Object Tracking

DGX agent

arXiv:2603.08199v2 Announce Type: replace Abstract: LiDAR-camera 3D multi-object tracking (MOT) combines rich visual semantics with accurate depth cues to improve trajectory consistency and tracking r

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

GDPevo: Evaluating Agent Self-Evolution on Real Business Tasks

DGX agent

arXiv:2608.03764v1 Announce Type: new Abstract: Agent self-evolution updates an agent's persistent state from prior experience and reuses it to solve related tasks more effectively. Evaluating self-ev

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

GENESIS: Towards Explainable Causal Discovery

DGX agent

arXiv:2608.03868v1 Announce Type: cross Abstract: Causal Discovery (CD) from observational data faces two fundamental challenges. First, purely statistical methods often lack the power to resolve stru

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Geo-Embed: Towards Unified Multimodal Embeddings for Urban Understanding

DGX agent

arXiv:2608.03826v1 Announce Type: new Abstract: Geospatial and urban applications increasingly require models to compare heterogeneous evidence across street-view imagery, remote-sensing observations,

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Getting the Parameters Right: A Difficulty-Graded Benchmark and Probe-Guided Training for LLM Tool Calls

DGX agent

arXiv:2608.03071v1 Announce Type: new Abstract: Large language model agents derive much of their capability from tool use. Existing research on tool use has largely focused on selecting the right tool

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

GoT-CD: Graph-of-Thoughts Causal Discovery and the Fragility of Post-hoc Path-Specific Fairness Audits

DGX agent

arXiv:2608.02877v1 Announce Type: new Abstract: Causal discovery recovers directed structure from observational data and is increasingly used in clinical settings to support mechanism reasoning and fa

model-releasesarxiv-cs-lg
5 Aug 2026
Model Releases

GUI-Lens: Coarse-to-Fine Cropping for GUI Grounding with General-Purpose VLMs

DGX agent

arXiv:2608.03270v1 Announce Type: cross Abstract: GUI grounding maps natural-language instructions to click locations and is essential for reliable GUI agents. The task remains difficult on high-resol

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

HomeSafeBench: A Benchmark for Embodied Vision-Language Models in Free-Exploration Home Safety Inspection

DGX agent

arXiv:2509.23690v2 Announce Type: replace-cross Abstract: Safety hazards in the home are a leading cause of preventable domestic injuries, motivating an automated inspector that actively explores a ho

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

How Closely Do LLM Reviews Align with Human Peer Review?

DGX agent

arXiv:2608.03659v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate scientific reviews, yet existing evaluations rarely examine whether different providers

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

HUKUKBERT: Domain-Specific Language Model for Turkish Law

DGX agent

arXiv:2604.04790v2 Announce Type: replace Abstract: Natural language processing (NLP) advances have powered a generation of LegalTech systems, but Turkish law remains under-served by domain-specific d

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

HyperFL: Query-Adaptive Representation Learning for Software Fault Localization

DGX agent

arXiv:2608.02967v1 Announce Type: cross Abstract: Software fault localization identifies the code locations responsible for reported issues and is a fundamental step toward automated debugging and pro

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

HyVIC: A Metric-Driven Spatio-Spectral Hyperspectral Image Compression Architecture Based on Variational Autoencoders

DGX agent

arXiv:2603.26468v2 Announce Type: replace Abstract: The rapid growth of hyperspectral data archives in remote sensing (RS) necessitates effective compression methods for storage and transmission. Rece

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

In-Context Collapse in Vision-Language Models and How to Mitigate it?

DGX agent

arXiv:2608.02830v1 Announce Type: cross Abstract: Many-shot in-context learning (ICL) lets vision-language models (VLMs) adapt from image--label demonstrations without weight updates, and is widely as

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

In-Context Pure Exploration in Continuous Decision Spaces

DGX agent

arXiv:2602.17976v2 Announce Type: replace-cross Abstract: In active sequential testing, also termed pure exploration, a learner is tasked with the goal to adaptively acquire information so as to ident

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Instruction Stacking Collapse: A Benchmark and the Capability-Dependent Value of Prompt Compilation

DGX agent

arXiv:2608.02639v1 Announce Type: cross Abstract: Production prompts rarely carry a single instruction. One system message may require valid JSON, a word limit, three citations, and a fixed tone at th

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Internalizing Academic Writing Workflows for Introduction Generation via Struct-Aware Policy Learning

DGX agent

arXiv:2608.03138v1 Announce Type: cross Abstract: Generating a rigorous paper introduction with large language models (LLMs) remains challenging, since it requires coordinating background, gap identif

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Intertemporal Preference Steering in Qwen3 via Contrastive Activation Addition

DGX agent

arXiv:2608.03892v1 Announce Type: new Abstract: We study linear representations of temporal horizon in the large language model Qwen3-32B and use them to change the model's time-related preferences, r

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Inverted Detection and Control in Steering Vectors

DGX agent

arXiv:2608.02957v1 Announce Type: new Abstract: Steering vectors (SVs) are widely used to influence the expression of concepts (e.g., truthfulness) in large language model outputs. A key assumption un

model-releasesarxiv-cs-lg
5 Aug 2026
← Previous
1…2627282930…357
Next →