Is there a simple solution to this?

Yes -- as long as your format doesn't have any more complexity that you haven't told us about yet.

There's a general approach to tree-building that's applicable here: Keep a stack of your "active" container, and every time you find a line with a "{" on the end of it, push a new container on to the stack. On lines with a "}", pop the active item off the stack. Every other piece of data that we find can get pushed into whichever container is currently active.

Of course, if your format includes escape characters, multi-line elements, or other complications, you may need to use one of the industrial-strength parser-generators... But for a simple format, we can roll our own.

Take a look at the output of the below, with and without $fewer_indents set true, and modify as desired.

sub parse_brackets { my @parse; my @stack = \@parse; my $fewer_indents = 1; # Try setting this to 0 or 1 my $line_no; foreach my $line ( @_ ) { $line_no ++; $line =~ s/\A\s+//; $line =~ s/\s+\Z//; if ( $line !~ /\S/ ) { next; } elsif ( $line =~ s/\s*\{$// ) { my @line = split ' ', $line; push @{ $stack[0] }, \@line; if ( $fewer_indents ) { unshift @stack, \@line; } else { my @kids = (); push @line, \@kids; unshift @stack, \@kids; } } elsif ( $line eq '}' ) { shift @stack; scalar @stack or die("Too many right brackets at line $line_no") +; } else { push @{ $stack[0] }, $line; } } return @parse; } use Data::Dumper; print Dumper( parse_brackets( split "\n", <<'EXAMPLE' ) ); page p1 { question 4B { label { Do you like your pie with ice cream? } single { 1 Yes 2 No } } question 4C { label { Do you like your pie with whipped cream? } single { 1 Yes 2 No } } } EXAMPLE

In reply to Re: Parsing a macro language by simonm
in thread Parsing a macro language by bluetrust

Title:
Use:  <p> text here (a paragraph) </p>
and:  <code> code here </code>
to format your post, it's "PerlMonks-approved HTML":



  • Posts are HTML formatted. Put <p> </p> tags around your paragraphs. Put <code> </code> tags around your code and data!
  • Titles consisting of a single word are discouraged, and in most cases are disallowed outright.
  • Read Where should I post X? if you're not absolutely sure you're posting in the right place.
  • Please read these before you post! —
  • Posts may use any of the Perl Monks Approved HTML tags:
    a, abbr, b, big, blockquote, br, caption, center, col, colgroup, dd, del, details, div, dl, dt, em, font, h1, h2, h3, h4, h5, h6, hr, i, ins, li, ol, p, pre, readmore, small, span, spoiler, strike, strong, sub, summary, sup, table, tbody, td, tfoot, th, thead, tr, tt, u, ul, wbr
  • You may need to use entities for some characters, as follows. (Exception: Within code tags, you can put the characters literally.)
            For:     Use:
    & &amp;
    < &lt;
    > &gt;
    [ &#91;
    ] &#93;
  • Link using PerlMonks shortcuts! What shortcuts can I use for linking?
  • See Writeup Formatting Tips and other pages linked from there for more info.