CLIP ViT-B/32
PulseAugur coverage of CLIP ViT-B/32 — every cluster mentioning CLIP ViT-B/32 across labs, papers, and developer communities, ranked by signal.
1 day(s) with sentiment data
-
Foundation models show promise for face attack detection but struggle with cross-dataset transfer
Researchers have investigated the effectiveness of foundation models for face presentation attack detection (PAD) by evaluating 24 frozen encoders using a unified linear-probing protocol. The study found that while thes…
-
Research: Feature alignment dictates multimodal fusion strategy
A new research paper proposes that feature alignment, rather than data scale, is the key factor in choosing between cross-attention and concatenation for multimodal fusion. The study demonstrates that when features are …
-
Paper challenges cosine similarity metric for neural representations
A new paper published on arXiv argues that mean-pooled cosine similarity, a common metric for comparing neural representations, is not length-invariant. The researchers demonstrate that sequence length alone can heavily…