Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Optimal hardware selection for AI-model learning
Luleå tekniska universitet, Institutionen för system- och rymdteknik.
2025 (Engelska)Självständigt arbete på avancerad nivå (yrkesexamen), 20 poäng / 30 hpStudentuppsats (Examensarbete)
Abstract [en]

Artificial Intelligence (AI) is a field of study that has seen rapid advancements in recent times, particularly due to the development of larger and more complicated models. The scale of recent models, in particular Large Language Models (LLMs), has led to a massive increase in computation required for development of these state-of-the-art models. Out of necessity the workload centered around these large-scale models is therefore oftentimes delegated to purpose-built computational clusters. These clusters are usually heterogenous when it comes to the types of deployed hardware which leads to increased difficulty in scheduling incoming jobs. This thesis was carried out in collaboration with Ericsson AB with the goal of developing a recommendation system for research clusters. Specifically, the system should be able to make accurate predictions about key running time metrics of AI training jobs using only the source code as input, which could help improve cluster efficiency. To this end, the system was broken down into several sequential components which are able to operate independently. The final system consists of a Transformer-CRF model for feature extraction of the source code followed by a Set-Transformer based architecture for regression of the features into a set of key metrics. These models are combined with algorithmic and mathematical solutions to form the purpose-built system. The mixed results on practical data will be put in context of the available data at hand along with training time constraints, forming a framework to carry the system into a production-ready state.

Ort, förlag, år, upplaga, sidor
2025. , s. 59
Nyckelord [en]
machine learning, artificial intelligence, transformer, conditional random fields, code analysis, running time estimation, regression
Nationell ämneskategori
Artificiell intelligens
Identifikatorer
URN: urn:nbn:se:ltu:diva-114430OAI: oai:DiVA.org:ltu-114430DiVA, id: diva2:1991789
Externt samarbete
Ericsson AB
Utbildningsprogram
Civilingenjör, Datateknik
Handledare
Examinatorer
Tillgänglig från: 2025-08-27 Skapad: 2025-08-25 Senast uppdaterad: 2025-10-21Bibliografiskt granskad

Open Access i DiVA

fulltext(839 kB)81 nedladdningar
Filinformation
Filnamn FULLTEXT02.pdfFilstorlek 839 kBChecksumma SHA-512
3bb091551e753bc7a9531a5d2c2cdc48aedbb6dade44a0a5cce1cbf2a0669e230db3cc8f09e681d281dacdb4a8f80342f565e2e5692d711cb232f4f1471c4f27
Typ fulltextMimetyp application/pdf

Sök vidare i DiVA

Av författaren/redaktören
Furhoff, Hannes
Av organisationen
Institutionen för system- och rymdteknik
Artificiell intelligens

Sök vidare utanför DiVA

GoogleGoogle Scholar
Totalt: 81 nedladdningar
Antalet nedladdningar är summan av nedladdningar för alla fulltexter. Det kan inkludera t.ex tidigare versioner som nu inte längre är tillgängliga.

urn-nbn

Altmetricpoäng

urn-nbn
Totalt: 171 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf