A Learnheuristic Algorithm Based on Thompson Sampling for the Heterogeneous and Dynamic Team Orienteering Problem
| dc.contributor.affiliation | Departamento de Estadística e Investigación Operativa Aplicadas y Calidad | |
| dc.contributor.affiliation | Centro de Investigación en Gestión e Ingeniería de Producción | |
| dc.contributor.affiliation | Escuela Politécnica Superior de Alcoy | |
| dc.contributor.author | Uguina, Antonio R. | es_ES |
| dc.contributor.author | Gomez, Juan F | es_ES |
| dc.contributor.author | Panadero, Javier | es_ES |
| dc.contributor.author | Martínez-Gavara, Anna | es_ES |
| dc.contributor.author | Juan, Angel A. | |
| dc.contributor.funder | European Commission | es_ES |
| dc.contributor.funder | Agencia Estatal de Investigación | es_ES |
| dc.contributor.funder | Ministerio de Ciencia e Innovación | es_ES |
| dc.date.accessioned | 2024-09-05T18:23:12Z | |
| dc.date.available | 2024-09-05T18:23:12Z | |
| dc.date.issued | 2024-06 | es_ES |
| dc.description.abstract | [EN] The team orienteering problem (TOP) is a well-studied optimization challenge in the field of Operations Research, where multiple vehicles aim to maximize the total collected rewards within a given time limit by visiting a subset of nodes in a network. With the goal of including dynamic and uncertain conditions inherent in real-world transportation scenarios, we introduce a novel dynamic variant of the TOP that considers real-time changes in environmental conditions affecting reward acquisition at each node. Specifically, we model the dynamic nature of environmental factors-such as traffic congestion, weather conditions, and battery level of each vehicle-to reflect their impact on the probability of obtaining the reward when visiting each type of node in a heterogeneous network. To address this problem, a learnheuristic optimization framework is proposed. It combines a metaheuristic algorithm with Thompson sampling to make informed decisions in dynamic environments. Furthermore, we conduct empirical experiments to assess the impact of varying reward probabilities on resource allocation and route planning within the context of this dynamic TOP, where nodes might offer a different reward behavior depending upon the environmental conditions. Our numerical results indicate that the proposed learnheuristic algorithm outperforms static approaches, achieving up to 25% better performance in highly dynamic scenarios. Our findings highlight the effectiveness of our approach in adapting to dynamic conditions and optimizing decision-making processes in transportation systems. | en_EN |
| dc.description.accrualMethod | S | es_ES |
| dc.description.bibliographicCitation | Uguina, AR.; Gomez, JF.; Panadero, J.; Martínez-Gavara, A.; Juan, AA. (2024). A Learnheuristic Algorithm Based on Thompson Sampling for the Heterogeneous and Dynamic Team Orienteering Problem. Mathematics. 12(11). https://doi.org/10.3390/math12111758 | es_ES |
| dc.description.issue | 11 | es_ES |
| dc.description.sponsorship | This work has been partially funded by the Spanish Ministry of Science and Innovation (PID2022-138860NB-I00, RED2022-134703-T) as well as by the SUN (HORIZON-CL4-2022-HUMAN01-14-101092612) and AIDEAS (HORIZON-CL4-2021-TWIN-TRANSITION-01-07-101057294) projects of the Horizon Europe program. | es_ES |
| dc.description.volume | 12 | es_ES |
| dc.identifier.doi | 10.3390/math12111758 | es_ES |
| dc.identifier.eissn | 2227-7390 | es_ES |
| dc.identifier.uri | https://riunet.upv.es/handle/10251/207466 | |
| dc.language | Inglés | es_ES |
| dc.publisher | MDPI AG | es_ES |
| dc.relation.ispartof | Mathematics | es_ES |
| dc.relation.pasarela | S\522326 | es_ES |
| dc.relation.projectID | info:eu-repo/grantAgreement/AEI/Plan Estatal de Investigación Científica y Técnica y de Innovación 2021-2023/PID2022-138860NB-I00/ES/INTELIGENCIA ARTIFICIAL E INTERNET DE LAS COSAS PARA OPTIMIZAR EL CONSUMO ENERGETICO EN EL TRANSPORTE CON VEHICULOS ELECTRICOS/ | es_ES |
| dc.relation.projectID | info:eu-repo/grantAgreement/EC/HE/101057294/EU/AI Driven industrial Equipment product life cycle boosting Agility, Sustainability and resilience/ | es_ES |
| dc.relation.projectID | info:eu-repo/grantAgreement/EC/HE/101092612/EU/Social and hUman ceNtered XR/ | es_ES |
| dc.relation.projectID | info:eu-repo/grantAgreement/MICINN//RED2022-134703-T/ | es_ES |
| dc.relation.publisherversion | https://doi.org/10.3390/math12111758 | es_ES |
| dc.rights | Reconocimiento (by) | es_ES |
| dc.rights.accessRights | Abierto | es_ES |
| dc.subject | Combinatorial optimization | es_ES |
| dc.subject | Team orienteering problem | es_ES |
| dc.subject | Reinforcement learning | es_ES |
| dc.subject | Learnheuristics | es_ES |
| dc.subject.classification | ESTADISTICA E INVESTIGACION OPERATIVA | es_ES |
| dc.title | A Learnheuristic Algorithm Based on Thompson Sampling for the Heterogeneous and Dynamic Team Orienteering Problem | es_ES |
| dc.type | Artículo | es_ES |
| dc.type.version | info:eu-repo/semantics/publishedVersion | es_ES |
| dspace.entity.type | Publication | |
| person.identifier | 490349 | |
| person.identifier.orcid | 0000-0003-1392-1776 | |
| relation.isAuthorOfPublication | 55e15b2b-1d12-4a12-b048-e805538d51e1 | |
| relation.isAuthorOfPublication.latestForDiscovery | 55e15b2b-1d12-4a12-b048-e805538d51e1 | |
| relation.isOrgUnitOfPublication | 73ebfca7-bf81-404f-861a-703ddec70645 | |
| relation.isOrgUnitOfPublication | 556fb9e2-3fb3-44c3-97e9-26873f979909 | |
| relation.isOrgUnitOfPublication | 96a57980-3fe4-46f5-8a98-7c3aac322fc5 | |
| relation.isOrgUnitOfPublication.latestForDiscovery | 73ebfca7-bf81-404f-861a-703ddec70645 | |
| upv.uuid | 8072e5fb-ebbf-4c77-8a73-4f1ed02142ce | es_ES |
Archivos
Bloque original
1 - 1 de 1
Cargando...
- Nombre:
- UguinaGomezPanadero - A Learnheuristic Algorithm Based on Thompson Sampling for the Heterogeneous....pdf
- Tamaño:
- 665.11 KB
- Formato:
- Adobe Portable Document Format
- Descripción:
- Versión editorial