ResearchResearch paperComputer Vision1 source · Oct 6, 2026

From the Drosophila Visual Connectome to General-Purpose Computer Vision

We develop ConnectomeX around FlyVision, a trainable architecture that preserves parallel ON/OFF processing, recurrent computation and population-level graph interaction while scaling model capacity across tasks.

Key points

  • Biological connectomes encode structured solutions to visual computation that may provide reusable inductive biases for artificial vision.
  • On ImageNet-1K, FlyVision Base and Large reached 60.79% and 66.25% top-1 accuracy with 1.8 and 3.7 million parameters, while a Large local-k7 model with a learned low-frequency branch reached 66.53%, compared with 69.25% for ResNet18 with 11.7 million parameters.
  • BrainAGE extends FlyVision to volumetric T1-weighted MRI by applying a shared ImageNet-pretrained FlyVision Large encoder to 24 sagittal, coronal and axial slices per scan and combining slice-level age estimates by confidence-modulated Gaussian voting.
  • These results show that a conserved connectome-informed computation can scale from compact recognition to large-scale natural and biomedical vision.

Sources (1)

  • [1]From the Drosophila Visual Connectome to General-Purpose Computer Vision
    arXiv (AI, ML, NLP, CV, robotics, multi-agent) · Oct 6, 02:22 PM
    We develop ConnectomeX around FlyVision, a trainable architecture that preserves parallel ON/OFF processing, recurrent computation and population-level graph interaction while scaling model capacity across tasks.
    Biological connectomes encode structured solutions to visual computation that may provide reusable inductive biases for artificial vision.

Extractive summary: sentences quoted from the sources.

Before this

  1. Oct 6, 2026Test-Time Adaptation of Quantized ViTs via Single-Pass Quantizer-Aligned Recalibration
  2. Oct 6, 2026Compact Robot Policies Need Fine-Grained Visual Representations
  3. Oct 6, 2026Two Halves are More than One: Phase-wise Velocity Distillation for Fast and High-Quality Image Generation
  4. Oct 6, 2026Later Is Better: Token Reduction for ViTs Under Distribution Shift
  5. Oct 1, 2026nvidia/PixelDiT2-ImageNet
  6. Sep 29, 2026Why Deep Learning Failed on Tables for a Decade - Frank Hutter

Related