Change search
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf
Djupinlärningsbaserad Monokulär 6D Pose Estimering som en Förbättrad Dockningsprocedur för Robotgräsklippare
Jönköping University, School of Engineering, JTH, Department of Computing.
2024 (Swedish)Independent thesis Advanced level (degree of Master (Two Years)), 20 credits / 30 HE creditsStudent thesisAlternative title
Deep Learning-Based Monocular 6D Pose Estimation as an Improved Robotic Lawnmower Docking Procedure (English)
Abstract [en]

This study investigates the integration of monocular CAD and RGB-based 6D pose estimation into robotic lawnmowers, focusing on charging station docking. A systematic literature review establishes the foundation by exploring methodologies, models, architectures, datasets, and relevant topics from existing literature on 6D pose estimation with CAD and RGB data. An approach for system development was developed integrating YOLOv7 and MegaPose, both deep learning-based, as the models for object detection and 6D pose prediction respectively. The detector model’s and 6D pose predictor’s performance across diverse datasets and distances highlights its potential, though specific challenges suggest the need for further refinement. A methodology for implementing an intelligent 6D pose estimation system for robotic lawnmower operation is developed through the Iterative Development Process where the results demonstrate promising outcomes. Notably, the detector model, trained exclusively on synthetically generated data, exhibits robust performance in detecting charging stations. The subsequent 6D pose predictor, MegaPose, demonstrates robust performance during neutral robotic lawnmower operation and scenarios with artificial degradation such as occlusion and water on the camera lens. Challenges encountered in real-world scenarios, such as low camera height, occlusions at close distances, and unclear imagery underscore the need for future improvements. This study bridges theoretical concepts with real-world implementation, laying a foundation for further advancements. The findings provide valuable insights for the robotic lawnmower domain and potentially for other robotic systems and domains by introducing them to or improving their capabilities via advanced autonomous navigation and interaction.

Place, publisher, year, edition, pages
2024. , p. 80
Keywords [en]
Artificial Intelligence, 6D Pose Estimation, Object Pose, Robotic Lawnmowers, Computer Vision, Deep Learning, Maneuvering, CAD, Monocular, RGB, Embedded Systems
National Category
Computer graphics and computer vision
Identifiers
URN: urn:nbn:se:hj:diva-65496OAI: oai:DiVA.org:hj-65496DiVA, id: diva2:1880158
External cooperation
Globe Technologies Sweden AB
Supervisors
Examiners
Available from: 2024-08-30 Created: 2024-06-30 Last updated: 2025-10-13Bibliographically approved

Open Access in DiVA

fulltext(12992 kB)907 downloads
File information
File name FULLTEXT01.pdfFile size 12992 kBChecksum SHA-512
02b658317579093b9b258ec0f80f3ca82908912c96fa11b142e0ad556f570bca7ac2c9568dc0f7e4477f292e0a4194458bb24796d1606b61d811073abe6bfa46
Type fulltextMimetype application/pdf

Search in DiVA

By author/editor
Appelberg, John
By organisation
JTH, Department of Computing
Computer graphics and computer vision

Search outside of DiVA

GoogleGoogle Scholar
Total: 911 downloads
The number of downloads is the sum of all downloads of full texts. It may include eg previous versions that are now no longer available

urn-nbn

Altmetric score

urn-nbn
Total: 648 hits
CiteExportLink to record
Permanent link

Direct link
Cite
Citation style
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Other style
More styles
Language
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Other locale
More languages
Output format
  • html
  • text
  • asciidoc
  • rtf