Endre søk
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Optimal hardware selection for AI-model learning
Luleå tekniska universitet, Institutionen för system- och rymdteknik.
2025 (engelsk)Independent thesis Advanced level (professional degree), 20 poäng / 30 hpOppgave
Abstract [en]

Artificial Intelligence (AI) is a field of study that has seen rapid advancements in recent times, particularly due to the development of larger and more complicated models. The scale of recent models, in particular Large Language Models (LLMs), has led to a massive increase in computation required for development of these state-of-the-art models. Out of necessity the workload centered around these large-scale models is therefore oftentimes delegated to purpose-built computational clusters. These clusters are usually heterogenous when it comes to the types of deployed hardware which leads to increased difficulty in scheduling incoming jobs. This thesis was carried out in collaboration with Ericsson AB with the goal of developing a recommendation system for research clusters. Specifically, the system should be able to make accurate predictions about key running time metrics of AI training jobs using only the source code as input, which could help improve cluster efficiency. To this end, the system was broken down into several sequential components which are able to operate independently. The final system consists of a Transformer-CRF model for feature extraction of the source code followed by a Set-Transformer based architecture for regression of the features into a set of key metrics. These models are combined with algorithmic and mathematical solutions to form the purpose-built system. The mixed results on practical data will be put in context of the available data at hand along with training time constraints, forming a framework to carry the system into a production-ready state.

sted, utgiver, år, opplag, sider
2025. , s. 59
Emneord [en]
machine learning, artificial intelligence, transformer, conditional random fields, code analysis, running time estimation, regression
HSV kategori
Identifikatorer
URN: urn:nbn:se:ltu:diva-114430OAI: oai:DiVA.org:ltu-114430DiVA, id: diva2:1991789
Eksternt samarbeid
Ericsson AB
Utdanningsprogram
Computer Science and Engineering, master's level
Veileder
Examiner
Tilgjengelig fra: 2025-08-27 Laget: 2025-08-25 Sist oppdatert: 2025-10-21bibliografisk kontrollert

Open Access i DiVA

fulltext(839 kB)81 nedlastinger
Filinformasjon
Fil FULLTEXT02.pdfFilstørrelse 839 kBChecksum SHA-512
3bb091551e753bc7a9531a5d2c2cdc48aedbb6dade44a0a5cce1cbf2a0669e230db3cc8f09e681d281dacdb4a8f80342f565e2e5692d711cb232f4f1471c4f27
Type fulltextMimetype application/pdf

Søk i DiVA

Av forfatter/redaktør
Furhoff, Hannes
Av organisasjonen

Søk utenfor DiVA

GoogleGoogle Scholar
Totalt: 81 nedlastinger
Antall nedlastinger er summen av alle nedlastinger av alle fulltekster. Det kan for eksempel være tidligere versjoner som er ikke lenger tilgjengelige

urn-nbn

Altmetric

urn-nbn
Totalt: 171 treff
RefereraExporteraLink to record
Permanent link

Direct link
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annet format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annet språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf