CViT Hybrid DeepFake Detector

Hybrid CNN + Vision Transformer model for detecting real vs fake images.

Architecture

  • EfficientNet-B0 (CNN)
  • Patch-based Transformer Encoder
  • Feature fusion head

Classes

  • Fake
  • Real

Input

  • RGB image
  • 224 × 224

Training

  • Optimizer: AdamW
  • LR: 1e-5
  • Epochs: 7
  • Dataset: DeepFake image dataset

Output

Binary classification (Fake / Real)

Downloads last month
6
Safetensors
Model size
7.27M params
Tensor type
F32
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support