International Peer-Reviewed JournalOpen AccessISSN 2456-8880
irejournals@gmail.com+91-7433024337

Home / Current Issue / Paper 1705287

1705287 Vol 7 · Issue 6 Download Paper

Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks

Archismita Ghosh Gaddam Prathik Kumar Paarth Prasad Dheeraj Kumar Samyak Jain Jatin Chopra

Subject area: Science,Engineering and Technology  ·  Area of research: Artificial intelligence

Abstract

This research paper presents a comprehensive exploration of generative artificial intelligence (AI) and its transformative impact on intelligent vehicles and vehicular networks. In the context of intelligent vehicles, the current state and future potential of generative AI technologies, emphasizing their applications in speech, audio, vision, and multimodal interactions are examined. The paper outlines critical future research areas, including domain adaptability, alignment, and multimodal integration, addressing associated ethical challenges. Simultaneously, recognizing the immense benefits of integrating generative AI into intelligent transportation systems, applications and challenges within vehicular networks are discussed. The integration of generative AI enhances various aspects, including navigation optimization, traffic prediction, and data generation, while facing challenges such as real-time processing and privacy concerns. To address these challenges, a multi-modality semantic-aware framework is proposed, leveraging text and image data to enhance generative AI service quality. A deep reinforcement learning (DRL)--based approach for resource allocation in generative AI-enabled vehicle-to-vehicle (V2V) communication is presented in a case study form. By synthesizing insights from both domains, this paper advocates for collaborative research efforts to unlock the full potential of generative AI, fostering transformative advancements in the driving experience and the evolution of intelligent vehicles and vehicular networks.

Keywords

Generative AI, Intelligent Vehicles, Vehicular Networks, Multi-modal Interactions, V2V, DRL

References

[1] C. Mou, X. Wang, L. Xie, Y. Wu, J. Zhang, Z. Qi, Y. Shan, and X. Qie, “T2I-Adapter: Learning adapters to dig out more controllable ability for text-to-image diffusion models,” 2023.

[2] Z. Yang, Y. Chai, D. Anguelov, Y. Zhou, P. Sun, D. Erhan, S. Rafferty, and H. Kretzschmar, “Surfelgan: Synthesizing realistic sensor data for autonomous driving,” in Proc. IEEE CVPR, June 2020.

[3] W. Jin, R. Barzilay, and T. Jaakkola, “Junction tree variational autoencoder for molecular graph generation,” in Proc. ICML, 2018. [Online]. Available: https://proceedings.mlr.press/v80/jin18a.html

[4] H. Du, Z. Li, D. Niyato, J. Kang, Z. Xiong, Xuemin, Shen, and D. I. Kim, “Enabling AI-generated content (AIGC) services in wireless edge networks,” 2023.

[5] S. W. Kim, J. Philion, A. Torralba, and S. Fidler, “DriveGAN: Towards a controllable high-quality neural simulation,” in Proc. IEEE CVPR, June 2021.

[6] M. Fisher, D. Ritchie, M. Savva, T. Funkhouser, and P. Hanrahan, “Example-based synthesis of 3D object arrangements,” ACM Trans. Graph., vol. 31, no. 6, 2012. [Online]. Available: https://doi.org/10. 1145/2366145.2366154

[7] E. A. van Dis, J. Bollen, W. Zuidema, R. van Rooij, and C. L. Bockting, “ChatGPT: Five priorities for research,” Nature, vol. 614, no. 7947, pp. 224–226, 2023.

[8] Y. Jing, Y. Yang, Z. Feng, J. Ye, Y. Yu, and M. Song, “Neural style transfer: A review,” IEEE Trans. Vis. Comput. Graph., vol. 26, no. 11, pp. 3365–3385, 2020.

[9] S. Kong and C. C. Fowlkes, “Recurrent pixel embedding for instance grouping,” in Proc. IEEE CVPR, June 2018.

[10] Lukas Stappen, Georgios Rizos, and Bjorn Schuller. X-aware: Context- ¨ aware human-environment attention fusion for driver gaze prediction in the wild. In Proceedings of the 2020 International Conference on Multimodal Interaction, ICMI ’20, pages 858—-867, New York, NY, USA, 2020. ACM.

[11] Lukas Stappen, Xinchen Du, Vincent Karas, Stefan Muller, and ¨ Bjorn W Schuller. Domain adaptation with joint learning for generic, ¨ optical car part recognition and detection systems (go-card). arXiv preprint arXiv:2006.08521, 2020.

[12] Sebastian Zepf, Javier Hernandez, Alexander Schmitt, Wolfgang Minker, and Rosalind W Picard. Driver emotion recognition for intelligent vehicles: A survey. ACM Computing Surveys (CSUR), 53(3):1–30, 2020.

[13] Simon Reiß, Alina Roitberg, Monica Haurilet, and Rainer Stiefelhagen. Deep classification-driven domain adaptation for cross-modal driver behavior recognition. In 2020 IEEE Intelligent Vehicles Symposium (IV), pages 1042–1047, 2020.

[14] Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, Andrew Tao, Jan Kautz, and Bryan Catanzaro. High-resolution image synthesis and semantic manipulation with conditional gans. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 8798– 8807, 2018.

[15] Mohamed Kari, Tobias Grosse-Puppendahl, Alexander Jagaciak, David Bethge, Reinhard Schutte, and Christian Holz. Soundsride: ¨ Affordance-synchronized music mixing for in-car audio augmented reality. In The 34th Annual ACM Symposium on User Interface Software and Technology, UIST ’21, pages 118–133, New York, NY, USA, 2021. ACM.

[16] Michael Braun, Anja Mainz, Ronee Chadowitz, Bastian Pfleging, and Florian Alt. At your service: Designing voice assistant personalities to improve automotive user interfaces. In Proceedings of the 2019 CHI Conference on Human Factors in Computing Systems, CHI ’19, page 1–11, New York, NY, USA, 2019. ACM.

[17] Robin Rombach, Andreas Blattmann, Dominik Lorenz, Patrick Esser, and Bjorn Ommer. High-resolution image synthesis with latent ¨ diffusion models. In Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition, pages 10684–10695, 2022.

[18] Ting-Chun Wang, Ming-Yu Liu, Jun-Yan Zhu, Andrew Tao, Jan Kautz, and Bryan Catanzaro. High-resolution image synthesis and semantic manipulation with conditional gans. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 8798– 8807, 2018.

[19] Phillip Isola, Jun-Yan Zhu, Tinghui Zhou, and Alexei A Efros. Image-to-image translation with conditional adversarial networks. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 1125–1134, 2017.

[20] Luciano Floridi and Massimo Chiriatti. Gpt-3: Its nature, scope, limits, and consequences. Minds and Machines, pages 1–14, 2020.

[21] Long Ouyang, Jeffrey Wu, Xu Jiang, Diogo Almeida, Carroll Wainwright, Pamela Mishkin, Chong Zhang, Sandhini Agarwal, Katarina Slama, Alex Ray, et al. Training language models to follow instructions with human feedback. Advances in Neural Information Processing Systems, 35:27730–27744, 2022.

[22] Christine Payne. Musenet. OpenAI Blog, 3, 2019.

[23] Prafulla Dhariwal, Heewoo Jun, Christine Payne, Jong Wook Kim, Alec Radford, and Ilya Sutskever. Jukebox: A generative model for music. arXiv preprint arXiv:2005.00341, 2020.

[24] Andrea Agostinelli, Timo I Denk, Zalan Borsos, Jesse Engel, Mauro ´ Verzetti, Antoine Caillon, Qingqing Huang, Aren Jansen, Adam Roberts, Marco Tagliasacchi, et al. Musiclm: Generating music from text. arXiv preprint arXiv:2301.11325, 2023.

[25] Tero Karras, Samuli Laine, Miika Aittala, Janne Hellsten, Jaakko Lehtinen, and Timo Aila. Analyzing and improving the image quality of stylegan. In Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition, pages 8110–8119, 2020.

[26] Pamela Mishkin, Lama Ahmad, Miles Brundage, Gretchen Krueger, and Girish Sastry. Dall·e 2 preview - risks and limitations. 2022.

[27] Zijie Guo, Rong Zhi, Wuqaing Zhang, Baofeng Wang, Zhijie Fang, Vitali Kaiser, Julian Wiederer, and Fabian Flohr. Generative model based data augmentation for special person classification. In 2020 IEEE Intelligent Vehicles Symposium (IV), pages 1675–1681, 2020.

[28] Ruben Villegas, Mohammad Babaeizadeh, Pieter-Jan Kindermans, Hernan Moraldo, Han Zhang, Mohammad Taghi Saffar, Santiago Castro, Julius Kunze, and Dumitru Erhan. Phenaki: Variable length video generation from open domain textual description. arXiv preprint arXiv:2210.02399, 2022.

[29] OpenAI. Gpt-4 technical report, 2023.

[30] Jonathan Shen, Ruoming Pang, Ron J Weiss, Mike Schuster, Navdeep Jaitly, Zongheng Yang, Zhifeng Chen, Yu Zhang, Yuxuan Wang, Rj Skerrv-Ryan, et al. Natural tts synthesis by conditioning wavenet on mel spectrogram predictions. In 2018 IEEE International Conference on Acoustics, Speech and Signal Processing (ICASSP), pages 4779– 4783. IEEE, 2018.

[31] Alec Radford, Jong Wook Kim, Tao Xu, Greg Brockman, Christine McLeavey, and Ilya Sutskever. Robust speech recognition via largescale weak supervision. arXiv preprint arXiv:2212.04356, 2022.

[32] Saleema Amershi, Maya Cakmak, William Bradley Knox, and Todd Kulesza. Power to the people: The role of humans in interactive machine learning. AI Magazine, 35(4):105–120, 2014.

[33] Advait Sarkar. Confidence, command, complexity: metamodels for structured interaction with machine intelligence. In PPIG, page 3, 2015.

[34] Taina Bucher. The algorithmic imaginary: Exploring the ordinary affects of facebook algorithms. Information, communication & society, 20(1):30–44, 2017.

[35] Malin Eiband, Hanna Schneider, Mark Bilandzic, Julian Fazekas-Con, Mareike Haug, and Heinrich Hussmann. Bringing transparency design into practice. In 23rd International Conference on Intelligent User Interfaces, pages 211–223, 2018.

How to cite this paper

Archismita Ghosh, Gaddam Prathik Kumar, Paarth Prasad, Dheeraj Kumar, Samyak Jain; Jatin Chopra "Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks" Iconic Research And Engineering Journals Volume 7 Issue 6 2023 Page 92-104
Archismita Ghosh, Gaddam Prathik Kumar, Paarth Prasad, Dheeraj Kumar, Samyak Jain; Jatin Chopra "Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks" Iconic Research And Engineering Journals, vol. 7, no. 6, Dec. 2023
Archismita Ghosh, Gaddam Prathik Kumar, Paarth Prasad, Dheeraj Kumar, Samyak Jain; Jatin Chopra (2023). Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks. Iconic Research And Engineering Journals, 7(6).
Archismita Ghosh, Gaddam Prathik Kumar, Paarth Prasad, Dheeraj Kumar, Samyak Jain; Jatin Chopra "Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks" Iconic Research And Engineering Journals, vol. 7, no. 6, Dec. 2023.
@article{1705287,
      author = {Archismita Ghosh, Gaddam Prathik Kumar, Paarth Prasad, Dheeraj Kumar, Samyak Jain; Jatin Chopra},
      title = {Synergizing Generative Intelligence: Advancements in Artificial Intelligence for Intelligent Vehicle Systems and Vehicular Networks},
      journal = {Iconic Research And Engineering Journals},
      year = {2023},
      volume = {7},
      number = {6},
      pages = {92-104},
      issn = {2456-8880},
      url = {https://www.irejournals.com/formatedpaper/1705287.pdf},
      abstract = {This research paper presents a comprehensive exploration of generative artificial intelligence (AI) and its transformative impact on intelligent vehicles and vehicular networks. In the context of intelligent vehicles, the current state and future potential of generative AI technologies, emphasizing their applications in speech, audio, vision, and multimodal interactions are examined. The paper outlines critical future research areas, including domain adaptability, alignment, and multimodal integration, addressing associated ethical challenges. Simultaneously, recognizing the immense benefits of integrating generative AI into intelligent transportation systems, applications and challenges within vehicular networks are discussed. The integration of generative AI enhances various aspects, including navigation optimization, traffic prediction, and data generation, while facing challenges such as real-time processing and privacy concerns. To address these challenges, a multi-modality semantic-aware framework is proposed, leveraging text and image data to enhance generative AI service quality. A deep reinforcement learning (DRL)--based approach for resource allocation in generative AI-enabled vehicle-to-vehicle (V2V) communication is presented in a case study form. By synthesizing insights from both domains, this paper advocates for collaborative research efforts to unlock the full potential of generative AI, fostering transformative advancements in the driving experience and the evolution of intelligent vehicles and vehicular networks.},
      keywords = {Generative AI, Intelligent Vehicles, Vehicular Networks, Multi-modal Interactions, V2V, DRL},
      month = {December},
  }