Models
Datasets
Spaces
Posts
Docs
Pricing
Log In
Sign Up

Collections

Discover the best community collections!

Collections including paper arxiv:2304.02643

Papers - SAM - Segment Anything Model

Prompt me a Dataset: An investigation of text-image prompting for historical image dataset creation using foundation models

Paper • 2309.01674 • Published Sep 4, 2023 • 2
Segment Anything

Paper • 2304.02643 • Published Apr 5, 2023 • 3
EgoLifter: Open-world 3D Segmentation for Egocentric Perception

Paper • 2403.18118 • Published Mar 26 • 10
A Multimodal Automated Interpretability Agent

Paper • 2404.14394 • Published Apr 22 • 20

Papers - Image - Segmentation

Image Segmentation using U-Net Architecture for Powder X-ray Diffraction Images

Paper • 2310.16186 • Published Oct 24, 2023 • 2
H-DenseUNet: Hybrid Densely Connected UNet for Liver and Tumor Segmentation from CT Volumes

Paper • 1709.07330 • Published Sep 21, 2017 • 2
Deep LOGISMOS: Deep Learning Graph-based 3D Segmentation of Pancreatic Tumors on CT scans

Paper • 1801.08599 • Published Jan 25, 2018 • 2
RTSeg: Real-time Semantic Segmentation Comparative Study

Paper • 1803.02758 • Published Mar 7, 2018 • 2

Papers - Image - Segment - Handwriting

Character Queries: A Transformer-based Approach to On-Line Handwritten Character Segmentation

Paper • 2309.03072 • Published Sep 6, 2023 • 2
Prompt me a Dataset: An investigation of text-image prompting for historical image dataset creation using foundation models

Paper • 2309.01674 • Published Sep 4, 2023 • 2
Segment Anything

Paper • 2304.02643 • Published Apr 5, 2023 • 3

Papers - Image - Handwriting Recognition

Data Incubation -- Synthesizing Missing Data for Handwriting Recognition

Paper • 2110.07040 • Published Oct 13, 2021 • 2
A Mixture of Expert Approach for Low-Cost Customization of Deep Neural Networks

Paper • 1811.00056 • Published Oct 31, 2018 • 2
Vulnerability Analysis of Transformer-based Optical Character Recognition to Adversarial Attacks

Paper • 2311.17128 • Published Nov 28, 2023 • 2
Data Generation for Post-OCR correction of Cyrillic handwriting

Paper • 2311.15896 • Published Nov 27, 2023 • 3

about 11 hours ago

FaceChain-SuDe: Building Derived Class to Inherit Category Attributes for One-shot Subject-Driven Generation

Paper • 2403.06775 • Published Mar 11 • 3
An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Paper • 2010.11929 • Published Oct 22, 2020 • 6
Data Incubation -- Synthesizing Missing Data for Handwriting Recognition

Paper • 2110.07040 • Published Oct 13, 2021 • 2
A Mixture of Expert Approach for Low-Cost Customization of Deep Neural Networks

Paper • 1811.00056 • Published Oct 31, 2018 • 2

Semantic Segmentation

Linking Points With Labels in 3D: A Review of Point Cloud Semantic Segmentation

Paper • 1908.08854 • Published Aug 23, 2019 • 1
Segment Anything

Paper • 2304.02643 • Published Apr 5, 2023 • 3
Segment and Caption Anything

Paper • 2312.00869 • Published Dec 1, 2023 • 18
Open-Vocabulary Audio-Visual Semantic Segmentation

Paper • 2407.21721 • Published Jul 31 • 8

Vision-Language

SILC: Improving Vision Language Pretraining with Self-Distillation

Paper • 2310.13355 • Published Oct 20, 2023 • 6
Woodpecker: Hallucination Correction for Multimodal Large Language Models

Paper • 2310.16045 • Published Oct 24, 2023 • 14
BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation

Paper • 2201.12086 • Published Jan 28, 2022 • 3
ImageNetVC: Zero-Shot Visual Commonsense Evaluation on 1000 ImageNet Categories

Paper • 2305.15028 • Published May 24, 2023 • 1

Attention Is All You Need

Paper • 1706.03762 • Published Jun 12, 2017 • 44
Retrieval-Augmented Generation for Knowledge-Intensive NLP Tasks

Paper • 2005.11401 • Published May 22, 2020 • 12
LoRA: Low-Rank Adaptation of Large Language Models

Paper • 2106.09685 • Published Jun 17, 2021 • 30
FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness

Paper • 2205.14135 • Published May 27, 2022 • 11

Company

© Hugging Face

TOS Privacy About Jobs

Website

Models Datasets Spaces Pricing Docs