hh.sePublikationer
Ändra sökning
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf
Detecting and Imputing Hidden Missing Values in Time Series Data: Case study: Alfa Laval
Högskolan i Halmstad, Akademin för informationsteknologi.
Högskolan i Halmstad, Akademin för informationsteknologi.
2024 (Engelska)Självständigt arbete på avancerad nivå (masterexamen), 20 poäng / 30 hpStudentuppsats (Examensarbete)
Abstract [en]

Although identifying missing values in regular time series is trivial,detecting them becomes a challenge with irregular timestamps. Toreduce the storage, our partner, Alfa Laval, uses an engineering trickto store measurements in time series databases only when their valuechanges. This solution, despite solving storage problems, can createproblems in data analysis. It also complicates the identification ofmissing values.

We address two problems: identifying hidden missing values fromirregular time series and developing effective imputation techniquesfor them. We use a rule-based approach to locate hidden missing val-ues tailored to the Alfa Laval dataset. Once we have identified the po-sition of hidden missing values, imputing them becomes the greaterchallenge, particularly when missing gaps are long. Our experimentsshow that while Linear Interpolation often outperforms LSTM andARIMA, it only creates a straight line between two points, failing tocapture the shape of the missing data. Consequently, in long-termgaps, we miss lots of informative fluctuations.

To address these limitations, we employ a pattern-based similar-ity search method, which effectively captures the value and shape oftime series data for more accurate imputation. This thesis presentsour novel approach, which we validate on a subset of Alfa Laval’ssensor data and three additional external datasets, demonstrating itsgeneralizability and effectiveness. While the rule-based identificationtechnique is particularly relevant to Alfa Laval’s data, our imputationtechnique serves as a general solution for time series imputation

Ort, förlag, år, upplaga, sidor
2024. , s. 112
Nyckelord [en]
hidden missing value, time series, pattern similarity search, similarity search, missing value imputation, time series
Nationell ämneskategori
Data- och informationsvetenskap
Identifikatorer
URN: urn:nbn:se:hh:diva-54251OAI: oai:DiVA.org:hh-54251DiVA, id: diva2:1882792
Externt samarbete
Alfa Laval
Handledare
Examinatorer
Tillgänglig från: 2024-07-16 Skapad: 2024-07-07 Senast uppdaterad: 2025-10-01Bibliografiskt granskad

Open Access i DiVA

fulltext(5058 kB)473 nedladdningar
Filinformation
Filnamn FULLTEXT02.pdfFilstorlek 5058 kBChecksumma SHA-512
ae783476bc1ff2a77e1c64430d0f058e59bb8dfb267d9f1bc7f2a38871173d36f273d54430c3aa04249bf6d78529b900d1cbeb0d9c4e3dd309ccd0ac9b2d5214
Typ fulltextMimetyp application/pdf

Av organisationen
Akademin för informationsteknologi
Data- och informationsvetenskap

Sök vidare utanför DiVA

GoogleGoogle Scholar
Totalt: 475 nedladdningar
Antalet nedladdningar är summan av nedladdningar för alla fulltexter. Det kan inkludera t.ex tidigare versioner som nu inte längre är tillgängliga.

urn-nbn

Altmetricpoäng

urn-nbn
Totalt: 683 träffar
RefereraExporteraLänk till posten
Permanent länk

Direktlänk
Referera
Referensformat
  • apa
  • ieee
  • modern-language-association-8th-edition
  • vancouver
  • Annat format
Fler format
Språk
  • de-DE
  • en-GB
  • en-US
  • fi-FI
  • nn-NO
  • nn-NB
  • sv-SE
  • Annat språk
Fler språk
Utmatningsformat
  • html
  • text
  • asciidoc
  • rtf