Downloads · 30 days
0
Girish666/ai_detect
ai_detect is a machine learning model from Girish666. Use it for the machine learning task on the model card, and read the license before you ship it in a product.
ICCV 2025 | 东南大学等提出LOTA:火眼金睛新范式,从“噪声”中秒速揪出AI生成图,准确率达98.9%
Downloads · 30 days
0
Access
Public
Updated Nov 6, 2025
Repo size
—
Likes
1
Public
Click a slice to open those files.
.py36 KB · 82%
From the Hugging Face model README
ICCV 2025 | 东南大学等提出LOTA:火眼金睛新范式,从“噪声”中秒速揪出AI生成图,准确率达98.9%
LOTA: Bit-Planes Guided AI-Generated Image Detection
(PS: 以上报道为自媒体编辑撰写,我们虽未向任何平台投稿,但欢迎翻译/宣传/讨论该工作,注明出处即可,谢谢)
Contributors: Hongsong Wang ([email protected]), Renxi Cheng ([email protected])
Southeast University, Nanjing, China
Novel solution for AI-generated image detection: We innovatively address AI-generated image detection based on bit-planes, and propose an efficient approach for noisy representation extraction.
Efficient pipeline design: We propose a simple yet effec tive pipeline with three modules: noise generation, patch selection and classification. We design a heuristic strategy called maximum gradient patch selection and introduce two effective classifiers: noise-based classifier and noise guided classifier. Our approach operates at millisecond level, nearly a hundred times faster than current methods.
Exceedingly superior performance: Extensive exper iments demonstrate the effectiveness of LOTA, which achieves 98.9% ACC onGenImage, showing great cross-generator generalization capability and outperforming existing mainstream methods by more than 11.9%.
We use GenImage for training and evaluation, which can be downloaded online. GenImage is composed of 8 subsets (BigGAN, Midjourney, Wukong, Stable_Diffusion_v1.4, Stable_Diffusion_v1.5, ADM, GLIDE, VQDM), each of which contains fake images and real images from ImageNet. Additionally, each subset are invided into training dataset and validating dataset, and we train LOTA on the training dataset of one subset (e.g., Stable_Diffusion_v1.5) and evaluate on the validating dataset of all subsets.
You can train the LOTA model on Stable_Diffusion_v1.5 of GenImage by running the following command:
python train.py --choice=[0, 0, 0, 0, 1, 0, 0, 0]
--image_root='Path/to/GenImage'
--save_path='Path/to/saved_weights'
--bit_mode='scaling'
--patch_size=32
--patch_mode='random'
You can evaluate LOTA on all subsets of GenImage by running the following command:
python test.py --choice=[1, 1, 1, 1, 1, 1, 1, 1]
--load='Path/to/saved_weights'
--image_root='Path/to/GenImage'
--bit_mode='scaling'
--patch_size=32
--patch_mode='max'
Additionally, we provide the pretrained weights trained on Stable_Diffusion_v1.4 (code: imjw) and Stable_Diffusion_v1.5 (code: a942) for evaluation. You can download the weights and easily evaluate LOTA on GenImage.
We provide the evaluation results in Results, which contains path_to_testing_images, true_label, predict_prob_real, and predict_prob_fake of all testing images in GenImage.
This repository borrows partially from CNNDetection, PatchCraft and SSP. Thanks for their work sincerely.
@InProceedings{Wang_2025_ICCV,
author = {Wang, Hongsong and Cheng, Renxi and Zhang, Yang and Han, Chaolei and Gui, Jie},
title = {LOTA: Bit-Planes Guided AI-Generated Image Detection},
booktitle = {Proceedings of the IEEE/CVF International Conference on Computer Vision (ICCV)},
month = {October},
year = {2025},
pages = {17246-17255}
}