AI Chat Paper
Note: Please note that the following content is generated by AMiner AI. SciOpen does not take any responsibility related to this content.
{{lang === 'zh_CN' ? '文章概述' : 'Summary'}}
{{lang === 'en_US' ? '中' : 'Eng'}}
Chat more with AI
Article Link
Collect
Submit Manuscript
Show Outline
Outline
Show full outline
Hide outline
Outline
Show full outline
Hide outline
Research Article | Open Access

Contrastive multi-view representation learning for multi-camera plant phenotyping: A cotton field study

Bio-Sensing, Automation, and Intelligence Laboratory, Department of Agricultural and Biological Engineering, Institute of Food and Agricultural Sciences, University of Florida, 1741 Museum Rd, Gainesville, 32603, Florida, USA
School of Computing, University of Georgia, 200 D. W. Brooks Drive, Georgia, Athens, 30602, USA
Show Author Information

Abstract

Attempts to deploy computer vision in agricultural tasks often suffer from a shortage of annotated data. One strategy to alleviate the impact of limited data is Self-Supervised Learning (SSL), which involves pre-training a model on a pretext task that utilizes automatically generated annotations. The primary objective of this study is to leverage a multi-camera view dataset of cotton boll images for contrastive learning in order to enable phenotyping tasks with minimal data annotation. This dataset was collected in the field using six camera views. The efficacy of two contrastive learning frameworks (SimCLR and MoCo) in producing representations when positive examples originate from different cameras was investigated, and a comprehensive study of how the camera positions affect performance was conducted. After self-supervised pre-training, linear evaluation and semi-supervised learning experiments were performed on boll detection and plot status downstream tasks. In general, using multiple camera views with SimCLR and MoCo improves cotton boll detection mean average precision by 14% compared to vanilla SimCLR and MoCo. Through careful investigation using synthetic data, it was determined that relative camera poses with an intermediate amount of overlap seem more likely to perform well. Neither MoCo nor SimCLR was consistently superior to the other in this context. The representations embed meaningful features about the cotton plants, such as overall boll density, but also less meaningful ones, such as lighting variations. This technique could potentially accelerate the development of phenotyping algorithms based on data collected from field robots.

References

【1】
【1】
 
 
Plant Phenomics
Article number: 100193

{{item.num}}

Comments on this article

Go to comment

< Back to all reports

Review Status: {{reviewData.commendedNum}} Commended , {{reviewData.revisionRequiredNum}} Revision Required , {{reviewData.notCommendedNum}} Not Commended Under Peer Review

Review Comment

Close
Close
Cite this article:
Petti D, Li C, Liu N. Contrastive multi-view representation learning for multi-camera plant phenotyping: A cotton field study. Plant Phenomics, 2026, 8(2): 100193. https://doi.org/10.1016/j.plaphe.2026.100193

10

Views

0

Crossref

0

Web of Science

0

Scopus

0

CSCD

Received: 02 September 2025
Revised: 09 February 2026
Accepted: 27 February 2026
Published: 06 March 2026
© 2026 The Authors. Nanjing Agricultural University.

This is an open access article under the CC BY license (http://creativecommons.org/licenses/by/4.0/).