aadex's picture
Add model card
fe37394 verified
|
Raw
History Blame Contribute Delete
846 Bytes
metadata
tags:
  - vision-transformer
  - image-classification
  - fire
  - imagenet100
license: apache-2.0

Fire Vit - IMAGENET100

This model was trained using the vit-analysis framework.

Model Details

  • Model Type: FIRE Vision Transformer
  • Dataset: imagenet100
  • Best Accuracy: 74.48%
  • Image Size: 224
  • Patch Size: 16
  • Hidden Dim: 192
  • Depth: 12
  • Num Heads: 3
  • MLP Dim: 768
  • Num Classes: 100

Training Configuration

  • Epochs: 120
  • Batch Size: 512
  • Learning Rate: 0.004
  • Weight Decay: 0.05
  • Label Smoothing: 0.1

Usage

import torch
from models import FIRESimpleVisionTransformer

# Load checkpoint
checkpoint = torch.load('fire_vit_imagenet100_best.pth')
model = ...  # Initialize model with same config
model.load_state_dict(checkpoint['state_dict'])