How to make textscan robust against non-matching lines?

Question

Joan Vazquez le 8 Avr 2021

0
Lien

Utiliser le lien direct vers cette question

https://fr.mathworks.com/matlabcentral/answers/796102-how-to-make-textscan-robust-against-non-matching-lines

Commenté : Stephen23 le 9 Avr 2021

data.txt

I have files with lines that I want to parse, preferably with textscan. In between those lines, there may be lines to be skipped (unpredictable format and abundance, but definetely new lines). What is the best way to deal with it?

E.g. for the data in attachment, this will stop outputiing #HELLOMATHWORKS messages after line 4.

fid = fopen('data.txt');
out = textscan(fid,'#HELLOMATHWORKS,%[^,],%n');
fclose(fid);

This is a MWE out of a large code base.

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Connectez-vous pour commenter.

Connectez-vous pour répondre à cette question.

Answer 1

Stephen23 le 8 Avr 2021

0
Lien

Utiliser le lien direct vers cette réponse

https://fr.mathworks.com/matlabcentral/answers/796102-how-to-make-textscan-robust-against-non-matching-lines#answer_670412

Modifié(e) : Stephen23 le 9 Avr 2021

Ouvrir dans MATLAB Online

data.txt

str = fileread('data.txt');
tkn = regexp(str,'#HELLOMATHWORKS,([^,]+),(\S+)','tokens');
tkn = vertcat(tkn{:})
tkn = 6×2 cell array
    {'COM1'}    {'2146'}
    {'COM1'}    {'2147'}
    {'COM1'}    {'2148'}
    {'COM1'}    {'2149'}
    {'COM1'}    {'2150'}
    {'COM1'}    {'2151'}
vec = str2double(tkn(:,2))
vec = 6×1
        2146
        2147
        2148
        2149
        2150
        2151

2 commentaires
Afficher AucuneMasquer Aucune

Joan Vazquez le 8 Avr 2021

Modifié(e) : Joan Vazquez le 8 Avr 2021

This does not produce the same output as my code:

tmp =

1×2 cell array

{6×1 cell} {6×1 double}

(Actually my messages have many more fields, this was just a MWE with 2... I have many similar functions using texscan to parse messages and I wanted to avoid refactoring them)

It is a good idea to work directly with regular expressions, but it seems that the formatSpec input parameter of textscan is not just any regular expression, it is more limited...

Anyway, It's OK for the moment, I'll accept the answer, thanks

Stephen23 le 9 Avr 2021

@Joan Vazquez: I presume that the text #HELLOMATHWORKS is not what is actually in your file. If the actual text contains some unique character that does not exist anywhere else in the file, you might be able to leverage the LineEnding/EndOfLine option to achieve the goal of reading the file data using textscan.

Connectez-vous pour commenter.

Answer 2

Joan Vazquez le 8 Avr 2021

1
Lien

Utiliser le lien direct vers cette réponse

https://fr.mathworks.com/matlabcentral/answers/796102-how-to-make-textscan-robust-against-non-matching-lines#answer_670177

Ouvrir dans MATLAB Online

This works, but it does not seem the best solution...Ideally, I would tell textscan "skip everything until a new line starts with #HELLOMATHWORKS"

filetext = fileread('data.txt');
expr = '[^\n]*#HELLOMATHWORKS[^\n]*';
% Find and return all lines that contain the text '#HELLOMATHWORKS'.
matches = regexp(filetext,expr,'match');
% Make it a 1xN char to feed textscan
goodlines = sprintf('%s\n', matches{:});
tmp = textscan(goodlines,'#HELLOMATHWORKS,%[^,],%n');

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Connectez-vous pour commenter.

How to make textscan robust against non-matching lines?

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Réponse acceptée

2 commentaires
Afficher AucuneMasquer Aucune

Plus de réponses (1)

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Voir également

Catégories

Tags

Produits

Version

Community Treasure Hunt

How to make textscan robust against non-matching lines?

0 commentaires Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Réponse acceptée

2 commentaires Afficher AucuneMasquer Aucune

Plus de réponses (1)

0 commentaires Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

Voir également

Catégories

Tags

Produits

Version

Community Treasure Hunt

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens

2 commentaires
Afficher AucuneMasquer Aucune

0 commentaires
Afficher -2 commentaires plus anciensMasquer -2 commentaires plus anciens