BARALDI, LORENZO
 Distribuzione geografica
Continente #
NA - Nord America 16.808
AS - Asia 12.368
EU - Europa 10.679
SA - Sud America 1.124
Continente sconosciuto - Info sul continente non disponibili 649
AF - Africa 201
OC - Oceania 68
Totale 41.897
Nazione #
US - Stati Uniti d'America 16.333
IT - Italia 4.897
SG - Singapore 3.403
CN - Cina 3.044
GB - Regno Unito 1.737
HK - Hong Kong 1.442
VN - Vietnam 1.086
DE - Germania 882
BR - Brasile 851
TR - Turchia 825
SE - Svezia 633
KR - Corea 516
FR - Francia 428
FI - Finlandia 365
RU - Federazione Russa 335
ID - Indonesia 316
JP - Giappone 313
BD - Bangladesh 311
IN - India 276
CA - Canada 274
NL - Olanda 249
UA - Ucraina 177
ES - Italia 164
MX - Messico 132
TW - Taiwan 121
AT - Austria 107
IE - Irlanda 104
MY - Malesia 100
BG - Bulgaria 93
TH - Thailandia 90
AR - Argentina 88
IQ - Iraq 82
BE - Belgio 77
PL - Polonia 73
CH - Svizzera 71
PH - Filippine 67
AU - Australia 59
ZA - Sudafrica 56
PK - Pakistan 51
RO - Romania 47
AE - Emirati Arabi Uniti 42
SA - Arabia Saudita 42
LT - Lituania 40
EC - Ecuador 39
GR - Grecia 39
CL - Cile 38
IL - Israele 38
DK - Danimarca 34
CO - Colombia 31
PT - Portogallo 31
VE - Venezuela 28
KE - Kenya 27
UZ - Uzbekistan 26
TN - Tunisia 23
MA - Marocco 21
DZ - Algeria 20
IR - Iran 20
JO - Giordania 20
NP - Nepal 20
EU - Europa 19
PE - Perù 19
CZ - Repubblica Ceca 18
KZ - Kazakistan 17
EG - Egitto 16
JM - Giamaica 16
AZ - Azerbaigian 13
PY - Paraguay 13
ET - Etiopia 10
SC - Seychelles 10
SY - Repubblica araba siriana 10
OM - Oman 9
UY - Uruguay 9
AL - Albania 8
BZ - Belize 8
CR - Costa Rica 8
HU - Ungheria 8
KH - Cambogia 8
LU - Lussemburgo 8
NZ - Nuova Zelanda 8
RS - Serbia 8
BH - Bahrain 7
HR - Croazia 7
MO - Macao, regione amministrativa speciale della Cina 7
SK - Slovacchia (Repubblica Slovacca) 7
BB - Barbados 6
BO - Bolivia 6
CY - Cipro 6
DO - Repubblica Dominicana 6
EE - Estonia 6
HN - Honduras 6
KG - Kirghizistan 6
MD - Moldavia 6
NO - Norvegia 6
KW - Kuwait 5
SN - Senegal 5
BA - Bosnia-Erzegovina 4
GT - Guatemala 4
LB - Libano 4
LK - Sri Lanka 4
LV - Lettonia 4
Totale 41.209
Città #
Singapore 2.152
Ashburn 1.670
Santa Clara 1.415
Fairfield 1.362
Hong Kong 1.207
Southend 958
Modena 914
San Jose 902
Hefei 859
Chandler 774
Elâzığ 670
Woodbridge 668
Seattle 614
Houston 596
Beijing 526
Cambridge 521
Wilmington 445
Ann Arbor 384
Los Angeles 361
Seoul 355
London 353
Ho Chi Minh City 342
Nyköping 342
Milan 329
Bologna 299
Jakarta 261
Council Bluffs 256
Dearborn 248
Hanoi 239
Jacksonville 237
Helsinki 236
New York 235
Buffalo 219
Rome 211
Chicago 208
The Dalles 206
Boardman 194
Tokyo 152
Reggio Emilia 142
San Diego 137
Lauterbourg 128
Munich 127
Parma 114
Nuremberg 100
Shanghai 97
Princeton 89
Sofia 87
Amsterdam 84
Bangkok 82
Dallas 82
São Paulo 82
Frankfurt am Main 79
Orem 79
Redwood City 79
Dublin 75
Florence 75
Montreal 74
Izmir 73
Kent 73
Phoenix 71
Dong Ket 69
Mexico City 64
Moscow 63
Naples 61
Salt Lake City 60
Eugene 59
Pisa 59
Manchester 57
Taipei 56
Chennai 55
Toronto 53
Bomporto 50
Bremen 49
Da Nang 49
Paris 48
Vienna 48
Atlanta 47
Haiphong 47
Kuala Selangor 46
Formigine 45
Warsaw 45
Falkenstein 44
Zurich 44
Brussels 43
Manila 42
Columbus 41
Turin 40
Piacenza 39
Fremont 36
Lappeenranta 35
Ottawa 35
Denver 33
Guangzhou 33
Johannesburg 33
Falls Church 32
San Francisco 32
Trento 32
Wilmette 32
Tampa 31
Düsseldorf 30
Totale 25.766
Nome #
Spaghetti Labeling: Directed Acyclic Graphs for Block-Based Connected Components Labeling 653
What was Monet seeing while painting? Translating artworks to photo-realistic images 651
MissRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models 637
Connected Components Labeling on DRAGs 601
Attentive Models in Vision: Computing Saliency Maps in the Deep Learning Era 589
Visual-Semantic Alignment Across Domains Using a Semi-Supervised Approach 562
Safe-CLIP: Removing NSFW Concepts from Vision-and-Language Models 527
Attentive Models in Vision: Computing Saliency Maps in the Deep Learning Era 525
Towards Cycle-Consistent Models for Text and Image Retrieval 501
Artpedia: A New Visual-Semantic Dataset with Visual and Contextual Sentences in the Artistic Domain 486
Modeling Multimodal Cues in a Deep Learning-based Framework for Emotion Recognition in the Wild 471
Automatic Image Cropping and Selection using Saliency: an Application to Historical Manuscripts 470
Connected Components Labeling on DRAGs: Implementation and Reproducibility Notes 463
YACCLAB - Yet Another Connected Components Labeling Benchmark 441
M-VAD Names: a Dataset for Video Captioning with Naming 433
Learning to Read L'Infinito: Handwritten Text Recognition with Synthetic Training Data 431
Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions 425
Explaining Digital Humanities by Aligning Images and Textual Descriptions 418
Layout analysis and content classification in digitized books 418
Aligning Text and Document Illustrations: towards Visually Explainable Digital Humanities 415
A Deep Multi-Level Network for Saliency Prediction 410
Watch Your Strokes: Improving Handwritten Text Recognition with Deformable Convolutions 410
Predicting Human Eye Fixations via an LSTM-based Saliency Attentive Model 396
A Browsing and Retrieval System for Broadcast Videos using Scene Detection and Automatic Annotation 392
A Hierarchical Quasi-Recurrent approach to Video Captioning 391
Recognizing social relationships from an egocentric vision perspective 388
Unveiling the Impact of Image Transformations on Deepfake Detection: An Experimental Analysis 382
Optimized Connected Components Labeling with Pixel Prediction 382
A Deep Siamese Network for Scene Detection in Broadcast Videos 380
A Video Library System Using Scene Detection and Automatic Tagging 378
Historical Document Digitization through Layout Analysis and Deep Content Classification 373
Image-to-Image Translation to Unfold the Reality of Artworks: an Empirical Analysis 373
Wiki-LLaVA: Hierarchical Retrieval-Augmented Generation for Multimodal LLMs 372
Context Change Detection for an Ultra-Low Power Low-Resolution Ego-Vision Imager 372
Hierarchical Boundary-Aware Neural Encoder for Video Captioning 371
Analysis and Re-use of Videos in Educational Digital Libraries with Automatic Scene Detection 370
Shot and Scene Detection via Hierarchical Clustering for Re-using Broadcast Video 364
Hand Segmentation for Gesture Recognition in EGO-Vision 363
Art2Real: Unfolding the Reality of Artworks via Semantically-Aware Image-to-Image Translation 362
Dual-Branch Collaborative Transformer for Virtual Try-On 359
Positive-Augmented Contrastive Learning for Image and Video Captioning Evaluation 358
Measuring scene detection performance 355
The Revolution of Multimodal Large Language Models: A Survey 354
Recognizing and Presenting the Storytelling Video Structure with Deep Multimodal Networks 354
SynthCap: Augmenting Transformers with Synthetic Data for Image Captioning 354
Ai4ar: An ai-based mobile application for the automatic generation of ar contents 353
Gesture Recognition using Wearable Vision Sensors to Enhance Visitors' Museum Experiences 352
Gesture Recognition in Ego-Centric Videos using Dense Trajectories and Hand Segmentation 346
SAM: Pushing the Limits of Saliency Prediction Models 342
From Show to Tell: A Survey on Deep Learning-based Image Captioning 341
Visual Saliency for Image Captioning in New Multimedia Services 339
Multi-Level Net: a Visual Saliency Prediction Model 334
LAMV: Learning to align and match videos with kernelized temporal layers 330
Explore and Explain: Self-supervised Navigation and Recounting 325
Tracing Information Flow in LLaMA Vision: A Step Toward Multimodal Understanding 324
A Novel Attention-based Aggregation Function to Combine Vision and Language 322
Multimodal Attention Networks for Low-Level Vision-and-Language Navigation 314
Paying More Attention to Saliency: Image Captioning with Saliency and Context Attention 306
Scene-driven Retrieval in Edited Videos using Aesthetic and Semantic Deep Features 304
Scene segmentation using temporal clustering for accessing and re-using broadcast video 303
Towards Video Captioning with Naming: a Novel Dataset and a Multi-Modal Approach 302
Embodied Agents for Efficient Exploration and Smart Scene Description 299
Meshed-Memory Transformer for Image Captioning 297
CaMEL: Mean Teacher Learning for Image Captioning 295
Recurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval 294
Towards Reliable Experiments on the Performance of Connected Components Labeling Algorithms 292
Boosting Modern and Historical Handwritten Text Recognition with Deformable Convolutions 292
Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities 286
A Unified Cycle-Consistent Neural Model for Text and Image Retrieval 277
A Computational Approach for Progressive Architecture Shrinkage in Action Recognition 277
Investigating Bidimensional Downsampling in Vision Transformer Models 271
NeuralStory: an Interactive Multimedia System for Video Indexing and Re-use 269
A Deep-learning-based approach to VM behavior Identification in Cloud Systems 269
Video action detection by learning graph-based spatio-temporal interactions 268
Retrieval-Augmented Transformer for Image Captioning 266
Adapt to Scarcity: Few-Shot Deepfake Detection via Low-Rank Adaptation 264
Intelligent Multimodal Artificial Agents that Talk and Express Emotions 254
Embodied Vision-and-Language Navigation with Dynamic Convolutional Filters 249
Semantically Conditioned Prompts for Visual Recognition under Missing Modality Scenarios 248
Embodied Navigation at the Art Gallery 245
Focus on Impact: Indoor Exploration with Intrinsic Motivation 244
Learning to Select: A Fully Attentive Approach for Novel Object Captioning 241
Hyperbolic Safety-Aware Vision-Language Models 240
Assessing the Role of Boundary-level Objectives in Indoor Semantic Segmentation 239
Revisiting The Evaluation of Class Activation Mapping for Explainability: A Novel Metric and Experimental Analysis 238
The Unreasonable Effectiveness of CLIP features for Image Captioning: an Experimental Analysis 237
Verifier Matters: Enhancing Inference-Time Scaling for Video Diffusion Models 235
Matching Faces and Attributes Between the Artistic and the Real Domain: the PersonArt Approach 235
Improving Indoor Semantic Segmentation with Boundary-level Objectives 234
Are Learnable Prompts the Right Way of Prompting? Adapting Vision-and-Language Models with Memory Optimization 233
RMS-Net: Regression and Masking for Soccer Event Spotting 233
BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues 230
FOSSIL: Free Open-Vocabulary Semantic Segmentation through Synthetic References Retrieval 226
Estimating (and fixing) the Effect of Face Obfuscation in Video Recognition 224
Towards Explainable Navigation and Recounting 220
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval 220
The LAM Dataset: A Novel Benchmark for Line-Level Handwritten Text Recognition 218
Multimodal Emotion Recognition in Conversation via Possible Speaker's Audio and Visual Sequence Selection 217
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering 214
With a Little Help from your own Past: Prototypical Memory Networks for Image Captioning 213
Totale 34.520
Categoria #
all - tutte 141.931
article - articoli 0
book - libri 0
conference - conferenze 0
curatela - curatele 0
other - altro 0
patent - brevetti 0
selected - selezionate 0
volume - volumi 0
Totale 141.931


Totale Lug Ago Sett Ott Nov Dic Gen Feb Mar Apr Mag Giu
2021/20223.237 183 139 240 169 84 214 186 230 311 335 825 321
2022/20232.767 364 299 246 227 301 283 92 226 386 75 144 124
2023/20242.525 243 170 246 312 439 164 120 183 70 201 131 246
2024/20258.439 710 242 260 496 1.229 925 512 650 1.081 567 803 964
2025/202615.647 1.084 780 1.263 1.528 2.134 1.059 1.930 1.297 1.295 1.683 883 711
2026/2027520 520 0 0 0 0 0 0 0 0 0 0 0
Totale 41.897