BARALDI, LORENZO
 Distribuzione geografica
Continente #
NA - Nord America 18.305
AS - Asia 12.470
EU - Europa 10.832
SA - Sud America 1.132
Continente sconosciuto - Info sul continente non disponibili 659
AF - Africa 203
OC - Oceania 69
Totale 43.670
Nazione #
US - Stati Uniti d'America 17.779
IT - Italia 5.010
SG - Singapore 3.418
CN - Cina 3.064
GB - Regno Unito 1.744
HK - Hong Kong 1.453
VN - Vietnam 1.091
DE - Germania 887
BR - Brasile 854
TR - Turchia 829
SE - Svezia 633
KR - Corea 521
FR - Francia 430
FI - Finlandia 366
RU - Federazione Russa 335
BD - Bangladesh 330
JP - Giappone 320
ID - Indonesia 318
CA - Canada 294
IN - India 282
NL - Olanda 254
UA - Ucraina 178
ES - Italia 166
MX - Messico 136
TW - Taiwan 121
AT - Austria 107
IE - Irlanda 104
MY - Malesia 100
BG - Bulgaria 93
TH - Thailandia 90
AR - Argentina 88
IQ - Iraq 82
BE - Belgio 78
CH - Svizzera 78
PL - Polonia 76
PH - Filippine 68
AU - Australia 60
ZA - Sudafrica 56
PK - Pakistan 52
RO - Romania 47
SA - Arabia Saudita 45
AE - Emirati Arabi Uniti 43
EC - Ecuador 40
LT - Lituania 40
GR - Grecia 39
CL - Cile 38
IL - Israele 38
DK - Danimarca 34
CO - Colombia 33
PT - Portogallo 32
VE - Venezuela 30
KE - Kenya 27
UZ - Uzbekistan 26
TN - Tunisia 23
DZ - Algeria 22
JM - Giamaica 21
MA - Marocco 21
IR - Iran 20
JO - Giordania 20
NP - Nepal 20
EU - Europa 19
PE - Perù 19
CZ - Repubblica Ceca 18
KZ - Kazakistan 17
EG - Egitto 16
AZ - Azerbaigian 13
CR - Costa Rica 13
PY - Paraguay 13
HN - Honduras 12
ET - Etiopia 10
SC - Seychelles 10
SY - Repubblica araba siriana 10
AL - Albania 9
OM - Oman 9
UY - Uruguay 9
BZ - Belize 8
HU - Ungheria 8
KH - Cambogia 8
LU - Lussemburgo 8
NZ - Nuova Zelanda 8
RS - Serbia 8
BH - Bahrain 7
HR - Croazia 7
KG - Kirghizistan 7
MO - Macao, regione amministrativa speciale della Cina 7
SK - Slovacchia (Repubblica Slovacca) 7
BB - Barbados 6
BO - Bolivia 6
CY - Cipro 6
DO - Repubblica Dominicana 6
EE - Estonia 6
GT - Guatemala 6
MD - Moldavia 6
NO - Norvegia 6
TT - Trinidad e Tobago 6
IS - Islanda 5
KW - Kuwait 5
SN - Senegal 5
BA - Bosnia-Erzegovina 4
LB - Libano 4
Totale 42.961
Città #
Singapore 2.163
Ashburn 1.744
Santa Clara 1.445
Fairfield 1.362
Hong Kong 1.213
San Jose 1.101
Southend 958
Modena 917
Hefei 859
Chandler 774
Elâzığ 670
Woodbridge 668
Seattle 617
Houston 598
Beijing 535
Cambridge 523
Council Bluffs 521
Wilmington 446
Ann Arbor 384
Los Angeles 374
Seoul 356
London 353
Ho Chi Minh City 343
Nyköping 342
Milan 333
Bologna 299
Jakarta 261
Dearborn 249
New York 244
Hanoi 239
Jacksonville 238
Helsinki 237
Buffalo 227
Rome 213
Chicago 211
The Dalles 208
Boardman 194
Tokyo 154
Reggio Emilia 142
San Diego 139
Lauterbourg 128
Munich 127
Parma 115
Nuremberg 100
Shanghai 98
Phoenix 94
Dallas 91
Princeton 89
Sofia 87
Amsterdam 84
Bangkok 82
São Paulo 82
Montreal 80
Frankfurt am Main 79
Orem 79
Redwood City 79
Florence 76
Dublin 75
Izmir 73
Kent 73
Columbus 72
Naples 70
Dong Ket 69
Mexico City 68
Moscow 63
Salt Lake City 62
Eugene 59
Pisa 59
Atlanta 57
Manchester 57
Taipei 56
Chennai 55
Toronto 54
Bomporto 50
Bremen 49
Da Nang 49
Paris 48
Vienna 48
Haiphong 47
Kuala Selangor 46
Formigine 45
Warsaw 45
Zurich 45
Falkenstein 44
Brussels 43
Manila 43
Turin 43
Piacenza 39
San Francisco 37
Fremont 36
Ottawa 36
Lappeenranta 35
North Charleston 35
Correggio 34
Denver 34
Guangzhou 33
Johannesburg 33
Falls Church 32
Trento 32
Wilmette 32
Totale 26.539
Nome #
Spaghetti Labeling: Directed Acyclic Graphs for Block-Based Connected Components Labeling 665
MissRAG: Addressing the Missing Modality Challenge in Multimodal Large Language Models 664
What was Monet seeing while painting? Translating artworks to photo-realistic images 657
Connected Components Labeling on DRAGs 611
Attentive Models in Vision: Computing Saliency Maps in the Deep Learning Era 599
From Show to Tell: A Survey on Deep Learning-based Image Captioning 575
Visual-Semantic Alignment Across Domains Using a Semi-Supervised Approach 566
Safe-CLIP: Removing NSFW Concepts from Vision-and-Language Models 535
Attentive Models in Vision: Computing Saliency Maps in the Deep Learning Era 533
Towards Cycle-Consistent Models for Text and Image Retrieval 507
Artpedia: A New Visual-Semantic Dataset with Visual and Contextual Sentences in the Artistic Domain 493
Connected Components Labeling on DRAGs: Implementation and Reproducibility Notes 483
Modeling Multimodal Cues in a Deep Learning-based Framework for Emotion Recognition in the Wild 476
Automatic Image Cropping and Selection using Saliency: an Application to Historical Manuscripts 474
YACCLAB - Yet Another Connected Components Labeling Benchmark 455
Learning to Read L'Infinito: Handwritten Text Recognition with Synthetic Training Data 442
Intelligent Multimodal Artificial Agents that Talk and Express Emotions 439
M-VAD Names: a Dataset for Video Captioning with Naming 438
Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions 434
Aligning Text and Document Illustrations: towards Visually Explainable Digital Humanities 427
Watch Your Strokes: Improving Handwritten Text Recognition with Deformable Convolutions 427
Explaining Digital Humanities by Aligning Images and Textual Descriptions 424
Layout analysis and content classification in digitized books 420
A Deep Multi-Level Network for Saliency Prediction 419
Predicting Human Eye Fixations via an LSTM-based Saliency Attentive Model 402
A Browsing and Retrieval System for Broadcast Videos using Scene Detection and Automatic Annotation 397
Recognizing social relationships from an egocentric vision perspective 394
A Hierarchical Quasi-Recurrent approach to Video Captioning 394
Unveiling the Impact of Image Transformations on Deepfake Detection: An Experimental Analysis 389
Optimized Connected Components Labeling with Pixel Prediction 389
A Deep Siamese Network for Scene Detection in Broadcast Videos 385
A Video Library System Using Scene Detection and Automatic Tagging 385
Wiki-LLaVA: Hierarchical Retrieval-Augmented Generation for Multimodal LLMs 382
Image-to-Image Translation to Unfold the Reality of Artworks: an Empirical Analysis 381
Hierarchical Boundary-Aware Neural Encoder for Video Captioning 380
Historical Document Digitization through Layout Analysis and Deep Content Classification 379
Context Change Detection for an Ultra-Low Power Low-Resolution Ego-Vision Imager 378
Analysis and Re-use of Videos in Educational Digital Libraries with Automatic Scene Detection 377
Hand Segmentation for Gesture Recognition in EGO-Vision 374
Shot and Scene Detection via Hierarchical Clustering for Re-using Broadcast Video 371
Art2Real: Unfolding the Reality of Artworks via Semantically-Aware Image-to-Image Translation 369
Dual-Branch Collaborative Transformer for Virtual Try-On 365
Recognizing and Presenting the Storytelling Video Structure with Deep Multimodal Networks 363
Positive-Augmented Contrastive Learning for Image and Video Captioning Evaluation 362
Ai4ar: An ai-based mobile application for the automatic generation of ar contents 360
SynthCap: Augmenting Transformers with Synthetic Data for Image Captioning 360
The Revolution of Multimodal Large Language Models: A Survey 359
Measuring scene detection performance 359
Gesture Recognition using Wearable Vision Sensors to Enhance Visitors' Museum Experiences 358
Gesture Recognition in Ego-Centric Videos using Dense Trajectories and Hand Segmentation 356
Visual Saliency for Image Captioning in New Multimedia Services 348
SAM: Pushing the Limits of Saliency Prediction Models 346
LAMV: Learning to align and match videos with kernelized temporal layers 339
Multi-Level Net: a Visual Saliency Prediction Model 338
Tracing Information Flow in LLaMA Vision: A Step Toward Multimodal Understanding 334
Explore and Explain: Self-supervised Navigation and Recounting 332
A Novel Attention-based Aggregation Function to Combine Vision and Language 331
Multimodal Attention Networks for Low-Level Vision-and-Language Navigation 321
Paying More Attention to Saliency: Image Captioning with Saliency and Context Attention 313
Scene segmentation using temporal clustering for accessing and re-using broadcast video 312
Embodied Agents for Efficient Exploration and Smart Scene Description 310
Towards Video Captioning with Naming: a Novel Dataset and a Multi-Modal Approach 308
Scene-driven Retrieval in Edited Videos using Aesthetic and Semantic Deep Features 308
Meshed-Memory Transformer for Image Captioning 306
Recurrence-Enhanced Vision-and-Language Transformers for Robust Multimodal Document Retrieval 304
CaMEL: Mean Teacher Learning for Image Captioning 300
Boosting Modern and Historical Handwritten Text Recognition with Deformable Convolutions 299
Towards Reliable Experiments on the Performance of Connected Components Labeling Algorithms 297
Contrasting Deepfakes Diffusion via Contrastive Learning and Global-Local Similarities 291
A Unified Cycle-Consistent Neural Model for Text and Image Retrieval 290
Investigating Bidimensional Downsampling in Vision Transformer Models 281
A Computational Approach for Progressive Architecture Shrinkage in Action Recognition 280
Adapt to Scarcity: Few-Shot Deepfake Detection via Low-Rank Adaptation 276
NeuralStory: an Interactive Multimedia System for Video Indexing and Re-use 276
Video action detection by learning graph-based spatio-temporal interactions 273
A Deep-learning-based approach to VM behavior Identification in Cloud Systems 271
Retrieval-Augmented Transformer for Image Captioning 269
Semantically Conditioned Prompts for Visual Recognition under Missing Modality Scenarios 255
Embodied Navigation at the Art Gallery 254
Embodied Vision-and-Language Navigation with Dynamic Convolutional Filters 253
Focus on Impact: Indoor Exploration with Intrinsic Motivation 248
Learning to Select: A Fully Attentive Approach for Novel Object Captioning 246
Hyperbolic Safety-Aware Vision-Language Models 246
Matching Faces and Attributes Between the Artistic and the Real Domain: the PersonArt Approach 245
Are Learnable Prompts the Right Way of Prompting? Adapting Vision-and-Language Models with Memory Optimization 244
Assessing the Role of Boundary-level Objectives in Indoor Semantic Segmentation 244
Revisiting The Evaluation of Class Activation Mapping for Explainability: A Novel Metric and Experimental Analysis 242
The Unreasonable Effectiveness of CLIP features for Image Captioning: an Experimental Analysis 242
Verifier Matters: Enhancing Inference-Time Scaling for Video Diffusion Models 241
Improving Indoor Semantic Segmentation with Boundary-level Objectives 238
RMS-Net: Regression and Masking for Soccer Event Spotting 236
BRIDGE: Bridging Gaps in Image Captioning Evaluation with Stronger Visual Cues 234
FOSSIL: Free Open-Vocabulary Semantic Segmentation through Synthetic References Retrieval 232
Estimating (and fixing) the Effect of Face Obfuscation in Video Recognition 228
Multimodal Emotion Recognition in Conversation via Possible Speaker's Audio and Visual Sequence Selection 227
The LAM Dataset: A Novel Benchmark for Line-Level Handwritten Text Recognition 225
Towards Explainable Navigation and Recounting 223
ALADIN: Distilling Fine-grained Alignment Scores for Efficient Image-Text Matching and Retrieval 223
Mapping High-level Semantic Regions in Indoor Environments without Object Recognition 221
Augmenting Multimodal LLMs with Self-Reflective Tokens for Knowledge-based Visual Question Answering 220
Totale 35.651
Categoria #
all - tutte 147.504
article - articoli 0
book - libri 0
conference - conferenze 0
curatela - curatele 0
other - altro 0
patent - brevetti 0
selected - selezionate 0
volume - volumi 0
Totale 147.504


Totale Lug Ago Sett Ott Nov Dic Gen Feb Mar Apr Mag Giu
2021/20222.915 0 0 240 169 84 214 186 230 311 335 825 321
2022/20232.767 364 299 246 227 301 283 92 226 386 75 144 124
2023/20242.525 243 170 246 312 439 164 120 183 70 201 131 246
2024/20258.439 710 242 260 496 1.229 925 512 650 1.081 567 803 964
2025/202615.647 1.084 780 1.263 1.528 2.134 1.059 1.930 1.297 1.295 1.683 883 711
2026/20272.293 635 1.637 21 0 0 0 0 0 0 0 0 0
Totale 43.670