arXiv is now an independent nonprofit! Learn more
License: arXiv.org perpetual non-exclusive license
arXiv:1809.03108v1 [cs.FL] 10 Sep 2018

Regular ω\omega-Languages with an Informative Right Congruence

Dana Angluin Email: angluin@cs.yale.edu Affiliation: Yale University    Dana Fisman Email: dana@cs.bgu.ac.il Affiliation: Ben-Gurion University
Abstract

A regular language is almost fully characterized by its right congruence relation. Indeed, a regular language can always be recognized by a dfa isomorphic to the automaton corresponding to its right congruence, henceforth the rightcon automaton. The same does not hold for regular ω\omega-languages. The right congruence of a regular ω\omega-language is not informative enough; many regular ω\omega-languages have a trivial right congruence, and in general it is not always possible to define an ω\omega-automaton recognizing a given language that is isomorphic to the rightcon automaton.

The class of weak regular ω\omega-languages does have an informative right congruence. That is, any weak regular ω\omega-language can always be recognized by a deterministic Bu¨\ddot{\textrm{u}}chi automaton that is isomorphic to the rightcon automaton. Weak regular ω\omega-languages reside in the lower levels of the expressiveness hierarchy of regular ω\omega-languages. Are there more expressive sub-classes of regular ω\omega-languages that have an informative right congruence? Can we fully characterize the class of languages with a trivial right congruence? In this paper we try to place some additional pieces of this big puzzle.

1 Introduction

Regular ω\omega-languages play a key role in reasoning about reactive systems. Algorithms for verification and synthesis of reactive system typically build on the theory of ω\omega–automata. The theory of ω\omega-automata enjoys many properties that the theory of automata on finite words enjoys. These make it amenable for providing the basis for analysis algorithms (e.g. model checking algorithms rely on the fact that emptiness can be checked in nondeterministic logarithmic space). However, in general, the theory of ω\omega-automata is much more involved than that of automata on finite words, and many fundamental questions, such as minimization, are still open.

One of the fundamental theorems of regular languages on finite words is the Myhill-Nerode theorem stating a one-to-one correspondence between the state of the minimal deterministic finite automaton (dfa) for a language LL and the equivalence classes of the right congruence of LL.11 1 Formal definitions are deferred to Section 2. When moving to ω\omega-words, there is no similar theorem, and there are many regular ω\omega-languages where any minimal automaton requires more states than the number of equivalence classes in the right congruence of the language. For instance, consider the ω\omega-language L=(a+b)aωL=(a+b)^{*}a^{\omega}. Its right congruence has only one equivalence class. That is, for any finite words xx and yy and any ω\omega-word ww we have that xwLxw\in L iff ywLyw\in L as membership in LL is determined only by the suffix. We say that the right congruence for LL is not informative enough.

The tight relationship between the equivalence classes of the right congruence and the states of a minimal dfa is at the heart of minimization and learning algorithms for regular languages of finite words, and seems to be a severe obstacle in obtaining efficient minimization and learning algorithms for regular ω\omega-languages. For this reason, we set ourselves to study classes of regular ω\omega-languages that do have a right congruence that is fully informative.

Several acceptance criteria are in use for ω\omega-automata, in particular, Bu¨\ddot{\textrm{u}}chi, co-Bu¨\ddot{\textrm{u}}chi, Muller and Parity. There are differences in the expressiveness of the corresponding deterministic automata. We use 𝔻𝔹\mathbb{DB}, 𝔻\mathbb{DC}, 𝔻\mathbb{DP} and 𝔻𝕄\mathbb{DM} to denote the classes of languages accepted by deterministic Bu¨\ddot{\textrm{u}}chi, co-Bu¨\ddot{\textrm{u}}chi, Muller and Parity automata, respectively. We also consider a version of Muller automata where acceptance is defined using transitions, refer to this acceptance criterion as Transition-Table, and use 𝔻𝕋\mathbb{DT} for the corresponding class of languages. The classes 𝔻𝕋\mathbb{DT}, 𝔻𝕄\mathbb{DM} and 𝔻\mathbb{DP} accept all regular ω\omega-languages whereas 𝔻𝔹\mathbb{DB} and 𝔻\mathbb{DC} are strictly less expressive and are dual to each other. The intersection of 𝔻𝔹\mathbb{DB} and 𝔻\mathbb{DC} is the class of weak regular ω\omega-languages. This class does have the property that any language in 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} has a fully informative right congruence. The regular ω\omega-languages can be arranged in an infinite hierarchy of expressive power as suggested by Wagner [34] and the class 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} corresponds to one of the lowest levels of the hierarchy.

We define the classes 𝕀𝔹\mathbb{IB}, 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM}, 𝕀𝕋\mathbb{IT} to be the class of regular ω\omega-languages that can be accepted by a Bu¨\ddot{\textrm{u}}chi, co-Bu¨\ddot{\textrm{u}}chi, Parity and Muller and Transition-Table automata, respectively, whose number of states equals the number of equivalence classes in the right congruence of the language. We show that these form a strictly inclusive hierarchy of expressiveness as shown in Fig. 1 on the left, and moreover in every class of the infinite Wagner hierarchy of regular ω\omega-languages there are languages whose right congruence is fully informative.

Noting another difficulty in inferring a regular ω\omega-language from examples of ω\omega-words in the language, we consider a further restriction on languages, that if uxωux^{\omega} is accepted by a minimal automaton for the language, then that automaton has a loop of size at most |x||x| in which uxωux^{\omega} loops. We term this property, being respective of the right congruence. This property is reminiscent of the property of being non-counting [14], and we show that a language that is non-counting is respective of its right congruence but the other direction does not necessarily hold. We define the classes 𝔹\mathbb{RB}, \mathbb{RC}, \mathbb{RP}, 𝕄\mathbb{RM}, 𝕋\mathbb{RT} of languages in 𝕀𝔹\mathbb{IB}, 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM}, 𝕀𝕋\mathbb{IT} that are respective of their right congruence. We show that these classes constitute a further restriction in terms of expressive power as shown in Fig. 1 on the right, and yet here as well, in every class of the infinite Wagner hierarchy of regular ω\omega-languages there are languages which are respective of their fully informative right congruence.

𝔻𝔹\mathbb{DB}\cap𝔻\mathbb{DC}𝕀𝔹\mathbb{IB}𝕀\mathbb{IC}𝕀\mathbb{IP}𝕀𝕄\mathbb{IM}𝕀𝕋\mathbb{IT} 𝕀𝔹\mathbb{IB}𝔹\mathbb{RB}𝕀\mathbb{IC}\mathbb{RC}𝕀\mathbb{IP}\mathbb{RP}𝕀𝕄\mathbb{IM}𝕄\mathbb{RM}𝕀𝕋\mathbb{IT}𝕋\mathbb{RT}

Figure 1: On the left, a summary of inclusion of the classes of ω\omega-automata that are isomorphic to their rightcon automaton. On the right a summary of inclusion of the classes of ω\omega-automata that are isomorphic to their rightcon automaton, as well as those that are in addition respective of their right congruence.

The rest of the paper is organized as follows. In Section 2 we provide the necessary definitions of the ω\omega-automata that we consider and of right congruence, state the well known results about their expressiveness, and briefly summarize the importance of the relation between the syntactic right congruence of a language and its minimal acceptor in learning algorithms. In Section 2 we state the relations between states of arbitrary ω\omega-automata for a regular ω\omega-language LL and its syntactic right congruence. In Section 3 we provide a full characterization of ω\omega-languages LL for which L\sim_{L} is trivial. In Section 4 we explore expressiveness results related to the classes of languages with an informative right congruence. In Section 5 we explore expressiveness results related to the classes of languages that not only have an informative right congruence, but are also respective of their right congruence. In Section 6 we explore closure properties of these classes, and in Section 7 we conclude. Due to lack of space some proofs are missing, these can be found in the full version of the paper at http://www.cis.upenn.edu/\simfisman/documents/AF_GandALF18_full.pdf

2 Preliminaries

An alphabet Σ\Sigma is a finite set of symbols. The set of finite words over Σ\Sigma is denoted by Σ\Sigma^{*}, and the set of infinite words, termed ω\omega-words, over Σ\Sigma is denoted by Σω\Sigma^{\omega}. We use ϵ\epsilon for the empty word and Σ+\Sigma^{+} for Σ{ϵ}\Sigma^{*}\setminus\{\epsilon\}. A language is a set of finite words, that is a subset of Σ\Sigma^{*}, while an ω\omega-language is a set of ω\omega-words, that is a subset of Σω\Sigma^{\omega}. For natural numbers ii and jj and a word ww, we use [i..j][i..j] for the set {i,i+1,,j}\{i,i+1,\ldots,j\}, w[i]w[i] for the ii-th letter of ww, and w[i..j]w[i..j] for the subword of ww starting at the ii-th letter and ending at the jj-th letter, inclusive.

Automata and Acceptors

An automaton is a tuple 𝒜=Σ,Q,λ,δ{{\mathpzc{A}}}=\langle\Sigma,Q,\lambda,\delta\rangle consisting of an alphabet Σ\Sigma, a finite set QQ of states, an initial state λQ\lambda\in Q, and a transition function δ:Q×Σ2Q{\delta:Q\times\Sigma\rightarrow 2^{Q}}. A run of an automaton on a finite word v=a1a2an{v=a_{1}a_{2}\ldots a_{n}} is a sequence of states q0,q1,,qn{q_{0},q_{1},\ldots,q_{n}} such that q0=λq_{0}=\lambda, and for each i0i\geq 0, qi+1δ(qi,ai+1){q_{i+1}\in\delta(q_{i},a_{i+1})}. A run on an infinite word is defined similarly and results in an infinite sequence of states. The transition function is naturally extended to a function from Q×ΣQ\times\Sigma^{*}, by defining δ(q,ϵ)=q\delta(q,\epsilon)=q and δ(q,av)=δ(δ(q,a),v){\delta(q,av)=\delta(\delta(q,a),v)} for qQ{q\in Q}, aΣ{a\in\Sigma} and vΣ{v\in\Sigma^{*}}. We often use 𝒜(v){{\mathpzc{A}}}(v) as a shorthand for δ(λ,v)\delta(\lambda,v) and |𝒜||{{\mathpzc{A}}}| for the number of states in QQ. We use 𝒜q{{\mathpzc{A}}}_{q} to denote the automaton Σ,Q,q,δ\langle\Sigma,Q,q,\delta\rangle obtained from 𝒜{{\mathpzc{A}}} by replacing the initial state with qq. We say that 𝒜{{\mathpzc{A}}} is deterministic if |δ(q,a)|1|\delta(q,a)|\leq 1 and complete if |δ(q,a)|1|\delta(q,a)|\geq 1, for every qQq\in Q and aΣa\in\Sigma.

By augmenting an automaton with an acceptance condition α\alpha, obtaining a tuple Σ,Q,λ,\langle\Sigma,Q,\lambda, δ,α\delta,\alpha\rangle, we get an acceptor, a machine that accepts some words and rejects others. An acceptor accepts a word if at least one of the runs on that word is accepting. For finite words the acceptance condition is a set FQF\subseteq Q and a run on a word vv is accepting if it ends in an accepting state, i.e., if δ(λ,v)\delta(\lambda,v) contains an element of FF. For infinite words, there are various acceptance conditions in the literature; we consider five: Bu¨\ddot{\textrm{u}}chi, co-Bu¨\ddot{\textrm{u}}chi, parity, Muller and Transition-Table [30]. The Bu¨\ddot{\textrm{u}}chi and co-Bu¨\ddot{\textrm{u}}chi acceptance conditions are also a set FQF\subseteq Q. A run of a Bu¨\ddot{\textrm{u}}chi automaton is accepting if it visits FF infinitely often. A run of a co-Bu¨\ddot{\textrm{u}}chi is accepting if it visits FF only finitely many times. A parity acceptance condition is a map κ:Q[1..k]\kappa:Q\rightarrow[1..k] (for some kk\in\mathbb{N}) assigning each state a color (or priority). A run is accepting if the minimal color visited infinitely often is odd. A Muller acceptance condition is a set of sets of states α={F1,F2,,Fk}\alpha=\{F_{1},F_{2},\ldots,F_{k}\} for some kk\in\mathbb{N} and FiQF_{i}\subseteq Q for i[1..k]i\in[1..k]. A run of a Muller automaton is accepting if the set SS of states visited infinitely often in the run is in α\alpha. A Transition-Table acceptance condition is a set α={T1,T2,,Tk}\alpha=\{T_{1},T_{2},\ldots,T_{k}\} of sets of transitions, where a transition is a tuple in Q×Σ×QQ\times\Sigma\times Q. A run of a Transition-Table automaton is accepting if the set TT of transitions visited infinitely often in the run is in α\alpha. We use 𝒜{{\llbracket{{\mathpzc{A}}}\rrbracket}} to denote the set of words accepted by a given acceptor 𝒜{{\mathpzc{A}}}.

:{\mathpzc{B}}:a{a}ab{ab}ϵ{\epsilon}a{a}Σ{a}\Sigma\setminus\{a\}Σ{b}\Sigma\setminus\{b\}bbΣ\Sigma𝒞:{\mathpzc{C}}:1{1}2{2}aabbbbaa:{\mathpzc{M}}:{{1},{2}}\hskip 13.14537pt\{\{1\},\{2\}\}1{1}2{2}bbaabbaa𝒫:{{{\mathpzc{P}}}}:κ(1)=2\kappa(1)=2κ(2)=1\kappa(2)=1κ(3)=0\kappa(3)=0κ(4)=0\kappa(4)=011223344aabbaabbaabbΣ\Sigma𝒯:{\mathpzc{T}}:{{(1,a,1)}}\{\{(1,a,1)\}\}11aabb

Figure 2: A dba {{{\mathpzc{B}}}} accepting (ΣaΣb)ω(\Sigma^{*}a\Sigma^{*}b)^{\omega} where Σ={a,b,c}\Sigma=\{a,b,c\}, a dca 𝒞{{{\mathpzc{C}}}} accepting (a+b)bω(a+b)^{*}b^{\omega}, a dma {{\mathpzc{M}}} accepting (a+b)bω(a+b)^{*}b^{\omega}, a dpa 𝒫{{{\mathpzc{P}}}} accepting (a+ba+bba)(aba)ω(a+ba+bba)^{*}(a^{*}ba)^{\omega}, and a dta 𝒯{{{\mathpzc{T}}}} accepting (a+b)aω(a+b)^{*}a^{\omega}.

We use three letter acronyms to describe acceptors, where the first letter is in {d,n}\{\textsc{d},\textsc{n}\} and denotes if the automaton is deterministic or nondeterministic. The second letter is one of {b,c,p,m,t}\{\textsc{b,c,p,m,t}\} for the first letter of the acceptance condition: Bu¨\ddot{\textrm{u}}chi, co-Bu¨\ddot{\textrm{u}}chi, parity, Muller or Transition-Table. The third letter is always a for acceptor. Figure 2 gives examples for a dba, dca, dpa, dma, and dta, and specifies the accepted languages. A language is said to be regular if it is accepted by a dfa. An ω\omega-language is said to be regular if it is accepted by a dma. An ω\omega-language is said to be weak if it is accepted by a dba as well as by a dca.

Complexity and expressiveness of sub-classes of regular ω\omega-languages

We use 𝔻𝔹\mathbb{DB}, 𝔻\mathbb{DC}, 𝔻\mathbb{DP}, 𝔻𝕄\mathbb{DM} and 𝔻𝕋\mathbb{DT} to denote the classes of languages accepted by dba, dca, dpa, dma and dta, respectively, and 𝔹\mathbb{NB}, \mathbb{NC}, \mathbb{NP}, 𝕄\mathbb{NM} and 𝕋\mathbb{NT} for the class of languages accepted by nba, nca, npa, nma and nta, respectively. The classes 𝔹\mathbb{NB}, \mathbb{NP}, 𝕄\mathbb{NM}, 𝕋\mathbb{NT}, 𝔻\mathbb{DP}, 𝔻𝕄\mathbb{DM} and 𝔻𝕋\mathbb{DT} are equi-expressive and contain all ω\omega-regular languages. The classes 𝔻𝔹\mathbb{DB} and 𝔻\mathbb{DC} are strictly less expressive and are dual to each other in the sense that L𝔻𝔹L\in\mathbb{DB} iff Lc𝔻L^{c}\in\mathbb{DC} where LcL^{c} is the complement language of LL, i.e. ΣωL\Sigma^{\omega}\setminus L. The classes \mathbb{NC} and 𝔻\mathbb{DC} are equi-expressive.

A subset SS of the automaton states, where for every s1,s2Ss_{1},s_{2}\in S there exists a string xΣ+x\in\Sigma^{+} such that δ(s1,x)=s2\delta(s_{1},x)=s_{2} is termed an SCC (abbreviating strongly connected component).22 2 Note that there is no requirement for SS to be maximal in this sense. A Muller automaton {{\mathpzc{M}}} can be seen as classifying its SCCs into accepting and rejecting. An important measure of the complexity of a Muller automaton is the number of alternations between accepting and rejecting SCCs along an inclusion chain. For instance, the Muller automaton {{\mathpzc{M}}} in Figure 2 has an inclusion chain of SCCs with one alternation: {1}\{1\} (accepting) {1,2}\subseteq\{1,2\} (rejecting); and the Muller automaton 𝒯{{{\mathpzc{T}}}} in Figure 3 whose only accepting SCC is {1,λ}\{1,\lambda\} has an inclusion chain of SCCs with two alternations: {1}\{1\} (rejecting) {1,λ}\subseteq\{1,\lambda\} (accepting) {1,λ,0}\subseteq\{1,\lambda,0\} (rejecting). Wagner [34] has shown that this complexity measure is language-specific and is invariant over all Muller automata accepting the same language. Under this view, 𝔻𝔹\mathbb{DB} is the class of languages where no superset of an accepting SCC can be rejecting, 𝔻\mathbb{DC} is the class of languages where no subset of an accepting SCC can be rejecting, and 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} is the class of languages where no alternation between accepting and rejecting SCCs is allowed along any inclusion chain. Thus, the language {{\llbracket{{\mathpzc{M}}}\rrbracket}} is not in 𝔻𝔹\mathbb{DB} and the language 𝒯{{\llbracket{{{\mathpzc{T}}}}\rrbracket}} is not in 𝔻𝔹𝔻\mathbb{DB}\cup\mathbb{DC}.

Right congruence and the rightcon automaton

An equivalence relation \sim on Σ\Sigma^{*} is a right congruence if xyx\sim y implies xvyvxv\sim yv for every x,y,vΣx,y,v\in\Sigma^{*}. The index of \sim, denoted ||{|\!\sim\!|} is the number of equivalence classes of \sim. Given a language LL its canonical right congruence L\sim_{L} is defined as follows: xLyx\sim_{L}y iff vΣ{\forall v\in\Sigma^{*}} we have xvLyvL{xv\in L}\iff{yv\in L}. For a word vΣv\in\Sigma^{*} the notation [v][v] is used for the class of \sim in which vv resides.

A right congruence \sim can be naturally associated with an automaton =Σ,Q,λ,δ{{\mathpzc{M}}}_{\sim}=\langle\Sigma,Q,\lambda,\delta\rangle as follows: the set of states QQ consists of the equivalence classes of \sim. The initial state λ\lambda is the equivalence class [ϵ][\epsilon]. The transition function δ\delta is defined by δ([u],a)=[ua]\delta([u],a)=[ua]. In the sequel, we use L{{\mathpzc{R}}}_{L} for the automaton MLM_{\sim_{L}} associated with the right congruence of a given language LL, and call it the rightcon automaton of LL.

Similarly, given an automaton =Σ,Q,λ,δ{{\mathpzc{M}}}=\langle\Sigma,Q,\lambda,\delta\rangle we can naturally associate with it a right congruence as follows: xyx\sim_{{\mathpzc{M}}}y iff (x)=(y){{\mathpzc{M}}}(x)={{\mathpzc{M}}}(y). The Myhill-Nerode Theorem states that a language LL is regular iff L\sim_{L} is of finite index. Moreover, if LL is accepted by a dfa 𝒜{{\mathpzc{A}}} then 𝒜\sim_{{\mathpzc{A}}} refines L\sim_{L}. Finally, the index of L\sim_{L} gives the size of the minimal dfa for LL.

For ω\omega-languages, the right congruence L\sim_{L} is defined similarly, by quantifying over ω\omega-words. That is, xLyx\sim_{L}y iff wΣω{\forall w\in\Sigma^{\omega}} we have xwLywL{xw\in L}\iff{yw\in L}. Given a deterministic automaton {{\mathpzc{M}}} we can define \sim_{{\mathpzc{M}}} exactly as for finite words. However, for ω\omega-regular languages, right congruence alone does not suffice to obtain a “Myhill-Nerode” characterization. As an example consider the language L=(a+b)(aba)ωL=(a+b)^{*}(aba)^{\omega}. We have that L\sim_{L} consists of just one equivalence class, since for any xΣx\in\Sigma^{*} and wΣωw\in\Sigma^{\omega} we have that xwLxw\in L iff ww has (aba)ω(aba)^{\omega} as a suffix. But an acceptor recognizing LL obviously needs more than a single state. Note that the other side of the story entails that there are ω\omega-automata that are minimal, although two different states recognize the same language. For instance, the dba {{{\mathpzc{B}}}} in Figure 2 is minimal for {{\llbracket{{{\mathpzc{B}}}}\rrbracket}} but ϵ=a=ab{{\llbracket{{{\mathpzc{B}}}}_{\epsilon}\rrbracket}}={{\llbracket{{{\mathpzc{B}}}}_{a}\rrbracket}}={{\llbracket{{{\mathpzc{B}}}}_{{ab}}\rrbracket}}, and the dma {{\mathpzc{M}}} of Figure 2 is minimal for {{\llbracket{{\mathpzc{M}}}\rrbracket}} but 1=2{{\llbracket{{\mathpzc{M}}}_{1}\rrbracket}}={{\llbracket{{\mathpzc{M}}}_{2}\rrbracket}}. In general the problem of minimizing dbas and dpas is known to be NP-complete [31].

Grammatical Inference

Grammatical inference or automata learning refers to the problem of designing algorithms for inferring an unknown language from good and bad examples, i.e. from words labeled by their membership in the language. The learning algorithm is required to return some concise representation of the language, typically an automaton. The task of a learning algorithm can thus be thought of as trying to distinguish the different necessary states of an automaton recognizing the language and establishing the transitions between them. For a regular language, by the Myhill-Nerode theorem, L\sim_{L} can be used to distinguish states. Indeed, if the algorithm learns that u1vLu_{1}v\in L and u2vLu_{2}v\notin L for some u1,u2,vΣu_{1},u_{2},v\in\Sigma^{*} then u1≁Lu2u_{1}\not\sim_{L}u_{2} and the words u1u_{1} and u2u_{2} must reach two different states of the minimal DFA for LL. Once all equivalence classes of L\sim_{L} are discovered, the automaton L{{\mathpzc{M}}}_{\sim_{L}} can be extracted, and by setting the state corresponding to the empty word to be the initial state, and states corresponding to positive examples as accepting, the minimal DFA is obtained. Many learning algorithms, e.g. Bierman and Feldman’s algorithm [10] for learning a regular language from a finite sample, and Angluin’s L{L^{*}} [3] algorithm for learning a regular language using membership and equivalence queries build on this idea.

The idea of trying to distinguish states using right congruence relations is in the essence of many learning algorithms for formalisms richer than regular language (c.f. [2, 11, 12, 15, 22, 26]). For instance, learning of deterministic weighted automata [8, 9] is founded on Fliess’s theorem [18] which is a generalization of the Myhill-Nerode theorem to the weighted automata setting.

For ω\omega-regular languages, learning algorithms encounter the problem that the right congruence is not informative enough. Maler and Pnueli [23] give a polynomial time algorithm for learning the class 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} using membership and equivalence queries. Their algorithm also works by trying to distinguish all equivalence classes of L\sim_{L} for the unknown language LL, and relies on the fact that L\sim_{L} is informative enough for the class 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC}. The problem of learning the full class of regular ω\omega-languages via membership and equivalence queries was considered open for many years [21]. It was then suggested by Farzan et al. [16] to use the reduction to finite words [13], building on the fact that two ω\omega-languages are equivalent iff they agree on the set of ultimately periodic words. We followed a different route [6, 7], and devised an algorithm for learning the full class of ω\omega-regular languages using families of DFAs as acceptors [4, 5]. Families of DFAs build on the notion of families of right congruences [24], a set of right congruence relations for a given regular ω\omega-language LL, which are enough to fully characterize LL. Both solutions, however, may encounter big automata in intermediate stages. The solution in [16] may involve DFAs of size 2n+22n2+n2^{n}+2^{2n^{2}+n} where nn is the number of states in a non-deterministic Bu¨\ddot{\textrm{u}}chi automaton for the language [13], and the solution in [6] may involve families of DFAs of size m2mm2^{m} where mm is the number of states in a minimal deterministic parity automaton for LL [17]. We do not go into further details here since they are not used in the current work, which focuses on languages for which the right congruence is informative enough (in hopes of providing a basis for more efficient learning algorithms for these restricted classes). The reader interested in these details is referred to [17].

Refinement of the right congruence

We show that as is the case in finitary regular languages, every deterministic ω\omega-automaton 𝒟{{{\mathpzc{D}}}} refines the rightcon automaton of the respective language 𝒟{{\llbracket{{{\mathpzc{D}}}}\rrbracket}}, and the automaton for the powerset construction of a given non-deterministic ω\omega-automaton 𝒩{{{\mathpzc{N}}}} refines the right congruence of the respective language 𝒩{{\llbracket{{{\mathpzc{N}}}}\rrbracket}}.

Proposition 1

Let 𝒟{{{\mathpzc{D}}}} be a deterministic ω\omega-automaton. Then 𝒟\sim_{{{{\mathpzc{D}}}}} refines 𝒟\sim_{{{\llbracket{{{\mathpzc{D}}}}\rrbracket}}}.

Let 𝒩{{{\mathpzc{N}}}} be a non-deterministic automaton. We use 𝒫𝒩{{{\mathpzc{P}}}}_{{{\mathpzc{N}}}} to denote the deterministic automaton obtained from 𝒩{{{\mathpzc{N}}}} by applying the powerset construction to 𝒩{{{\mathpzc{N}}}}. That is, if 𝒩=(Σ,Q,q0,δ,α){{{\mathpzc{N}}}}=(\Sigma,Q,q_{0},\delta,\alpha) then 𝒫𝒩=(Σ,2Q,{q0},δ){{{\mathpzc{P}}}}_{{{\mathpzc{N}}}}=(\Sigma,2^{Q},\{q_{0}\},\delta^{\prime}) where δ(S,a)=qSδ(q,a)\delta^{\prime}(S,a)=\cup_{q\in S}\delta(q,a) for any SQS\subseteq Q and aΣa\in\Sigma.

Proposition 2

Let 𝒩{{{\mathpzc{N}}}} be a non-deterministic ω\omega-automaton. Then 𝒫𝒩\sim_{{{{\mathpzc{P}}}}_{{{\mathpzc{N}}}}} refines 𝒩\sim_{{{\llbracket{{{\mathpzc{N}}}}\rrbracket}}}.

3 ω\omega-Languages with a trivial right congruence

If LL is a regular ω\omega-language such that |L|=1{|\sim_{L}|}=1, we say that the rightcon automaton of LL is trivial. In this case, the rightcon automaton conveys almost no information about LL. It was shown in [32] that there are 2202^{2^{\aleph_{0}}} ω\omega-languages for which the rightcon is trivial. In Proposition 3.5 we characterize those regular ω\omega-languages that have a trivial rightcon automaton.

If w1w_{1} and w2w_{2} are ω\omega-words, then w1w_{1} is a finite variant of w2w_{2}, denoted w1=w2w_{1}=_{\infty}w_{2}, if there exist finite words x1x_{1} and x2x_{2} and an ω\omega-word ww such that w1=x1ww_{1}=x_{1}w and w2=x2ww_{2}=x_{2}w. The following shows that languages with a trivial rightcon automaton ignore differences between finite variants.

Proposition 3

Let LL be a regular ω\omega-language such that |L|=1{|\sim_{L}|}=1. Let w1,w2Σωw_{1},w_{2}\in\Sigma^{\omega}. If w1=w2w_{1}=_{\infty}w_{2} then w1Lw_{1}\in L iff w2Lw_{2}\in L.

Proof 3.4.

Because w1=w2w_{1}=_{\infty}w_{2}, there exist x1x_{1}, x2x_{2} and ww such that w1=x1ww_{1}=x_{1}w and w2=x2ww_{2}=x_{2}w. Because |L|=1{|\sim_{L}|}=1, [x1]L=[x2]L[x_{1}]_{\sim_{L}}=[x_{2}]_{\sim_{L}}. Thus x1wLx_{1}w\in L iff x2wLx_{2}w\in L, so w1Lw_{1}\in L iff w2Lw_{2}\in L.

Clearly if L=ΣvωL=\Sigma^{*}v^{\omega} for some vΣv\in\Sigma^{*} then its rightcon automaton is trivial. Does this hold in general when L=ΣLL=\Sigma^{*}L^{\prime} for some ω\omega-regular language LL^{\prime}? The following example shows that the answer is negative. Let Σ={a,b}\Sigma=\{a,b\} and L1=ΣLL_{1}=\Sigma^{*}L^{\prime} for L=(abω+baω)L^{\prime}=(ab^{\omega}+ba^{\omega}) then L1\sim_{L_{1}} has four equivalence classes.

However, if L=ΣRωL=\Sigma^{*}R^{\omega} for some regular language RR we can show that then |L|=1|\sim_{L}|=1. Is this also a necessary condition for having a trivial right congruence? The answer again is negative. For example, the language L2=(a+b)(aω+bωCLOSEL_{2}=(a+b)^{*}(a^{\omega}+b^{\omega}) has |L2|=1{|\sim_{L_{2}}|}=1, but it is not of the form (a+b)Rω(a+b)^{*}R^{\omega} for any regular set RR.

The following proposition provides a full characterization of regular ω\omega-language with a trivial rightcon automaton.

Proposition 3.5.

A regular ω\omega-language has a trivial rightcon automaton iff L=Σ(R1ω+R2ω++Rkω)L=\Sigma^{*}(R_{1}^{\omega}+R_{2}^{\omega}+\ldots+R_{k}^{\omega}) for some regular languages R1,,RkR_{1},\ldots,R_{k}.

Proof 3.6.

Let =(Σ,Q,λ,δ,α){\mathpzc{M}}=(\Sigma,Q,\lambda,\delta,\alpha) be a DMA for LL with no unreachable states. Let α={S1,S2,,Sk}\alpha=\{S_{1},S_{2},\ldots,S_{k}\}. We can assume wlog that all SiS_{i}’s are strongly connected (because an SiS_{i} that is not an SCC can be omitted). For i[1..k]i\in[1..k] let sis_{i} be some state in SiS_{i} and let RiR_{i} be the regular set of finite words that traverse {\mathpzc{M}} starting at sis_{i}, ending in sis_{i}, and visiting all states in SiS_{i} and no other states.

Assume that |L|=1{|\sim_{L}|}=1. Let L=Σ(R1ω+R2ω++Rkω)L^{\prime}=\Sigma^{*}(R_{1}^{\omega}+R_{2}^{\omega}+\ldots+R_{k}^{\omega}). We claim that L=LL=L^{\prime}. Let wLw\in L. Then w=xww=xw^{\prime}, where for some ii, xΣx\in\Sigma^{*} reaches siSis_{i}\in S_{i} and wRiωw^{\prime}\in R_{i}^{\omega}, so wLw\in L^{\prime}. Conversely, if wLw\in L^{\prime} then w=xww=xw^{\prime}, where xΣx\in\Sigma^{*} and for some ii, wRiωw^{\prime}\in R_{i}^{\omega}. Because MM contains no unreachable states, there exists yΣy\in\Sigma^{*} such that yy reaches state siSis_{i}\in S_{i}. Then ywLyw^{\prime}\in L, so by Proposition 3, xwLxw^{\prime}\in L.

For the converse, suppose that R1,,RkR_{1},\ldots,R_{k} are regular languages and L=Σ(R1ω++Rkω)L=\Sigma^{*}(R_{1}^{\omega}+\ldots+R_{k}^{\omega}). If |L|>1{|\sim_{L}|}>1 then there exist x,yΣx,y\in\Sigma^{*} such that [x]L[y]L[x]_{L}\neq[y]_{L}, so wlog assume w[x]Lw\in[x]_{L} and w[y]Lw\not\in[y]_{L}. Because xwLxw\in L, there exists an ii such that xwΣRiωxw\in\Sigma^{*}R_{i}^{\omega}. Thus for some xΣx^{\prime}\in\Sigma^{*} and elements w1,w2,w_{1},w_{2},\ldots of RiR_{i}, xw=xw1w2xw=x^{\prime}w_{1}w_{2}\cdots. Hence there exists x′′Σx^{\prime\prime}\in\Sigma^{*} and w′′Riωw^{\prime\prime}\in R_{i}^{\omega} such that xw=xx′′w′′xw=xx^{\prime\prime}w^{\prime\prime}. But then yw=yx′′w′′yw=yx^{\prime\prime}w^{\prime\prime}, and yx′′w′′ΣRiωyx^{\prime\prime}w^{\prime\prime}\in\Sigma^{*}R_{i}^{\omega}, which implies that ywLyw\in L, a contradiction. Thus we must have |L|=1{|\sim_{L}|}=1.

4 ω\omega-Languages with an informative right congruence

We turn to examine the cases where the right congruence is as informative as it can be; that is the rightcon automaton is isomorphic to an ω\omega-automaton recognizing the respective language. We use 𝕀𝔹\mathbb{IB} (resp. 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM}, 𝕀𝕋\mathbb{IT}) to denote the class of languages for which the rightcon automaton L{{\mathpzc{R}}}_{L} is isomorphic to a dba (resp. dca, dpa, dma, dta) accepting the language LL.

A small experiment

We were curious to see what are the odds that a randomly generated Muller automaton will be isomorphic to its rightcon automaton, i.e. fully informative. We ran a small experiment in which we generated a random Muller automaton over an alphabet of cardinality 3, with 2 accepting strongly connected sets, and tested whether it turned out to be isomorphic to its rightcon automaton. The procedure was to try to distinguish states of the random dma using 100,000 random ultimately periodic ω\omega-words. If all states were successfully distinguished then the dma is certainly isomorphic to its rightcon automaton, and was declared as such. If we failed to distinguish at least 2 states, we declared the dma as non-isomorphic, though it might be that more tests would distinguish the undistinguished states and the dma may in fact be isomorphic. So the probability of a randomly generated dma being isomorphic to its rightcon automaton may be higher than what is suggested by our results.

We generated dmas with 5, 6, 7, 8, 9, and 10 states; 100 of each size. The results are summarized in the following table. We find it interesting that in most of the cases a randomly generated dma turns out to be isomorphic to its rightcon automaton, suggesting that this property is not rare. We defer a more careful study of the extent to which random automata have informative right congruences for further research.

Number of states 5 6 7 8 9 10
Isomorphic 85 93 88 96 96 94
Not Isomorphic 15 7 12 4 4 6
:{\mathpzc{M}}:{{λ,1}}\{\{\lambda,1\}\}𝒫:{\mathpzc{P}}:κ(0)=0\kappa(0)=0κ(λ)=1\kappa(\lambda)=1κ(1)=2\kappa(1)=2λ\lambda001100110,10,10011{{{\mathpzc{B}}}} or 𝒞:{{{\mathpzc{C}}}}:11223300aabbcca,b,ca,b,cb,cb,caabbccaa:{{1},{2}}{{\mathpzc{M}}}^{\prime}:\{\{1\},\{2\}\}1{1}2{2}3{3}aabbccbbccaaa,b,ca,b,c𝒯:{{(1,a,1)},{(1,b,1)}}{{{\mathpzc{T}}}}:\{\{(1,a,1)\},\{(1,b,1)\}\}1{1}aabb
Figure 3: A dma {{\mathpzc{M}}} and an equivalent dpa 𝒫{{{\mathpzc{P}}}} for a language LPML_{PM} such that LPM𝕀𝕄𝔻𝔹𝔻L_{PM}\in\mathbb{IM}\setminus\mathbb{DB}\cap\mathbb{DC} and LPM𝕀𝔻𝔹𝔻L_{PM}\in\mathbb{IP}\setminus\mathbb{DB}\cap\mathbb{DC}. Second from the left, when regarded as a dba {{{\mathpzc{B}}}} we have 𝕀𝔹𝔻{{\llbracket{{{\mathpzc{B}}}}\rrbracket}}\in\mathbb{IB}\setminus\mathbb{DC}. When regarded as a dca 𝒞{{{\mathpzc{C}}}} we have 𝒞𝕀𝔻𝔹{{\llbracket{{{\mathpzc{C}}}}\rrbracket}}\in\mathbb{IC}\setminus\mathbb{DB}. A dma {{\mathpzc{M}}}^{\prime} such that 𝕀𝕄𝕀{{\llbracket{{\mathpzc{M}}}^{\prime}\rrbracket}}\in\mathbb{IM}\setminus\mathbb{IP}, and a dta 𝒯{{{\mathpzc{T}}}} such that 𝒯𝕋𝕄{{\llbracket{{{\mathpzc{T}}}}\rrbracket}}\in\mathbb{RT}\setminus\mathbb{RM}.

Expressiveness results

As mentioned earlier, all weak regular ω\omega-languages, i.e. all languages that are in 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} are isomorphic to their right congruence. We turn to the question of whether there exist languages outside this class that are isomorphic to their right congruence.

Staiger [32] has shown that 𝔻𝔹𝔻𝕀𝕄\mathbb{DB}\cap\mathbb{DC}\subseteq\mathbb{IM}. It is easy to see that this entails 𝔻𝔹𝔻𝕀\mathbb{DB}\cap\mathbb{DC}\subseteq\mathbb{IP}. We show that both inclusions are strict.

Proposition 4.7.

𝔻𝔹𝔻𝕀𝕄\mathbb{DB}\cap\mathbb{DC}\subsetneq\mathbb{IM} and 𝔻𝔹𝔻𝕀\mathbb{DB}\cap\mathbb{DC}\subsetneq\mathbb{IP}

Proof 4.8.

In [32, Thm. 24] Staiger showed that 𝔻𝔹𝔻𝕀𝕄\mathbb{DB}\cap\mathbb{DC}\subseteq\mathbb{IM}. From this, since any dba can be recognized by an isomorphic dpa (by setting the accepting states color 11 and the non-accepting states color 22) and a dca can be recognized by an isomorphic dpa (by setting the FF states color 00 and the non-FF states color 11) it follows that 𝔻𝔹𝔻𝕀\mathbb{DB}\cap\mathbb{DC}\subseteq\mathbb{IP}.

To show that the inclusion is strict we show that the language LPML_{PM} recognized by the automaton in Fig. 3 on the left, either when regarded as a dma {{\mathpzc{M}}} or as a dpa 𝒫{{{\mathpzc{P}}}} is in 𝕀𝕄𝕀\mathbb{IM}\cap\mathbb{IP} but not in 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC}. One can verify that \sim_{{\llbracket{{\mathpzc{M}}}\rrbracket}} is isomorphic to {{\mathpzc{M}}}. (Note that (0011)ω(0011)^{\omega} and (0110)ω(0110)^{\omega} are sufficient to distinguish {λ,0,1}\{\lambda,0,1\}.) As mentioned in the preliminaries, this language is not in 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} since it has alternation between accepting and rejecting SCCs along the inclusion chain {1}{1,λ}{1,λ,0}\{1\}\subseteq\{1,\lambda\}\subseteq\{1,\lambda,0\}.

It is easy to see that both 𝕀𝕄\mathbb{IM} and 𝕀\mathbb{IP} subsume both 𝕀𝔹\mathbb{IB} and 𝕀\mathbb{IC}. The same example used in the proof of Proposition 4.7 can be used to show that the inclusion is strict.

Proposition 4.9.

𝕀𝔹𝕀𝕀𝕄\mathbb{IB}\cup\mathbb{IC}\subsetneq\mathbb{IM} and 𝕀𝔹𝕀𝕀\mathbb{IB}\cup\mathbb{IC}\subsetneq\mathbb{IP}

It is thus interesting to see whether there are any languages in 𝕀𝔹\mathbb{IB} (or 𝕀\mathbb{IC}) that are not already in 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC}. The answer is affirmative.

Proposition 4.10.

𝔻𝔹𝔻𝕀𝔹\mathbb{DB}\cap\mathbb{DC}\subsetneq\mathbb{IB} and 𝔻𝔹𝔻𝕀\mathbb{DB}\cap\mathbb{DC}\subsetneq\mathbb{IC}

Proof 4.11.

Assume L𝔻𝔹𝔻L\in\mathbb{DB}\cap\mathbb{DC}. By Proposition 4.7 we have L𝕀𝕄L\in\mathbb{IM}. Suppose {{\mathpzc{M}}} is a minimal dma for LL. It follows, as explained in the preliminaries, that in {{\mathpzc{M}}} there is no alternation between accepting and rejecting SCCs along any inclusion chain. Therefore, defining a dba {{{\mathpzc{B}}}}_{{\mathpzc{M}}} from {{\mathpzc{M}}} by changing the acceptance condition to {q|qS\{q~|~q\in S for some accepting SCC S}S\} gives ={\llbracket{{{\mathpzc{B}}}}_{{\mathpzc{M}}}\rrbracket}={\llbracket{{\mathpzc{M}}}\rrbracket} and thus L𝕀𝔹L\in\mathbb{IB}. Similarly, defining a dca 𝒞{{{\mathpzc{C}}}}_{{\mathpzc{M}}} from {{\mathpzc{M}}} by changing the acceptance condition to {q|qS\{q~|~q\in S for some rejecting SCC S}S\} gives 𝒞={\llbracket{{{\mathpzc{C}}}}_{{\mathpzc{M}}}\rrbracket}={\llbracket{{\mathpzc{M}}}\rrbracket} and thus L𝕀L\in\mathbb{IC}. This completes the inclusion part.

To see that the inclusion is strict for 𝕀𝔹\mathbb{IB} consider the dba {{{\mathpzc{B}}}} from Fig. 3. By [20, Lemma 2] for every language LL recognized by a dba {{{\mathpzc{B}}}}, if LL is also in 𝔻\mathbb{DC}, then a dca embodied in the structure of {{{\mathpzc{B}}}} can be defined. Since none of the dcas embodied in {{{\mathpzc{B}}}} accepts the same language as {\llbracket{{{\mathpzc{B}}}}\rrbracket}, it follows that 𝔻𝔹𝔻{\llbracket{{{\mathpzc{B}}}}\rrbracket}\in\mathbb{DB}\setminus\mathbb{DC}. Thus 𝔻𝔹𝔻{{{\mathpzc{B}}}}\notin\mathbb{DB}\cap\mathbb{DC}.

Next we show that \sim_{{\llbracket{{{\mathpzc{B}}}}\rrbracket}} has at least 4 equivalence classes. The ω\omega-word (ababc)ω(ababc)^{\omega} is accepted only from states 00 and 33 and the ω\omega-word (babca)ω(babca)^{\omega} is accepted only from states 00 and 11; these two experiments distinguish all 4 states. Since by Proposition 1 {{{\mathpzc{B}}}} refines \sim_{{\llbracket{{{\mathpzc{B}}}}\rrbracket}}, it follows that the automaton for \sim_{{\llbracket{{{\mathpzc{B}}}}\rrbracket}} is isomorphic to {{{\mathpzc{B}}}}, thus 𝕀𝔹{\llbracket{{{\mathpzc{B}}}}\rrbracket}\in\mathbb{IB}.

The proof for strictness for 𝕀\mathbb{IC} is dual, using the dca 𝒞{{{\mathpzc{C}}}} in Fig. 3 and [20, Lemma 2] which states also that for every language LL recognized by a dca 𝒞{{{\mathpzc{C}}}} if LL is also in 𝔻𝔹\mathbb{DB} then a dba embodied in the structure of 𝒞{{{\mathpzc{C}}}} can be defined.

Since 𝕀𝔹𝔻𝔹\mathbb{IB}\subseteq\mathbb{DB} and 𝕀𝔻\mathbb{IC}\subseteq\mathbb{DC}, from 𝔻𝔹𝔻𝕀𝔹𝕀\mathbb{DB}\cap\mathbb{DC}\subseteq\mathbb{IB}\cap\mathbb{IC} we get that 𝔻𝔹𝔻=𝕀𝔹𝕀\mathbb{DB}\cap\mathbb{DC}=\mathbb{IB}\cap\mathbb{IC}.

Corollary 4.12.

𝔻𝔹𝔻=𝕀𝔹𝕀\mathbb{DB}\cap\mathbb{DC}=\mathbb{IB}\cap\mathbb{IC}

It is shown in [33] that any dma can be defined on a dta with the same structure, and that the converse is not true. For instance, the language (a+b)aω(a+b)^{*}a^{\omega} can be defined by the one-state dta 𝒯{{{\mathpzc{T}}}} in Fig. 2, but no dma with one state accepts it. This shows 𝕀𝕄\mathbb{IM} is strictly contained in 𝕀𝕋\mathbb{IT}.

Proposition 4.13 ([33]).

𝕀𝕄𝕀𝕋\mathbb{IM}\subsetneq\mathbb{IT}

The last missing part of the puzzle of the inclusions of the subsets 𝕀𝔹\mathbb{IB}, 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM}, 𝕀𝕋\mathbb{IT} is provided in the following proposition showing that 𝕀\mathbb{IP} is strictly contained in 𝕀𝕄\mathbb{IM}.

Proposition 4.14.

𝕀𝕀𝕄\mathbb{IP}\subsetneq\mathbb{IM}

Proof 4.15.

Inclusion follows since any dpa can be converted into a dma on the same structure (by setting an SCC to accepting iff the minimal color in it is odd). We claim that the dma {{\mathpzc{M}}}^{\prime} in Figure 3 is in 𝕀𝕄𝕀\mathbb{IM}\setminus\mathbb{IP}. To see that it is in 𝕀𝕄\mathbb{IM}, note that caωca^{\omega} distinguishes state 22 from the other states, and aωa^{\omega} distinguishes states 11 and 33. Thus the rightcon automaton has 33 states, and since by Proposition 1, {{\mathpzc{M}}}^{\prime} refines {{\mathpzc{R}}}_{{{\llbracket{{\mathpzc{M}}}^{\prime}\rrbracket}}} we get that they are isomorphic. To see that it is not in 𝕀\mathbb{IP}, note that to define a dpa on the same structure we need to give state 11 an odd color, so that when the set of states visited inf. often is {1}\{1\} it will accept. For the same reason we need to give state 22 an odd color. But then, when the set of states visited infinitely often is {1,2}\{1,2\}, the automaton will accept as well, while it needs to reject.

These relations are summarized in Figure 1 on the left. The last question is then how complex can a language in 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM}, or 𝕀𝕋\mathbb{IT} be? We show that such languages can be arbitrarily complex. That is, for every class 𝔻𝕄n,mp\mathbb{DM}_{n,m}^{p} of the Wagner Hierarchy, there is a language L𝕀𝕄𝕀𝕀𝕋L\in\mathbb{IM}\cap\mathbb{IP}\cap\mathbb{IT} that is in 𝔻𝕄n,mp\mathbb{DM}_{n,m}^{p} and not in any proper subclass of the Wagner Hierarchy. Thus, 𝕀𝕋\mathbb{IT}, 𝕀𝕄\mathbb{IM} and 𝕀\mathbb{IP} include classes as complex as can be33 3 A similar result is mentioned without a proof in a footnote in [24], with credit to “N. Gutleben (personal communication)”. , as measured by the Wagner Hierarchy.44 4 Due to lack of space we do not include the full definition of the Wagner Hierarchy. For details, we refer the reader to [34],[27, Chapter V].

Proposition 4.16.

Let n,mn,m be two natural numbers and p{+,,±}p\in\{+,-,\pm\}. Let 𝔻𝕄n,mp\mathbb{DM}_{n,m}^{p} denote the Wagner Hierarchy class with a maximum of nn alternations in each inclusion chain starting with polariy pp, and a sequence of at most mm chains of alternating polarity. Then exists a language LΣωL\in\Sigma^{\omega} for Σ={a,b}\Sigma=\{a,b\} such that L𝕀𝕋𝕀𝕄𝕀𝔻𝕄n,mpL\in\mathbb{IT}\cap\mathbb{IM}\cap\mathbb{IP}\cap\mathbb{DM}_{n,m}^{p} and L𝔻𝕄n1,mpL\notin\mathbb{DM}_{n\!-\!1,m}^{p} and L𝔻𝕄n,m1pL\notin\mathbb{DM}_{n,m\!-\!1}^{p}.

𝒟n,m+:{{{{{{\mathpzc{D}}}}_{n,m}^{+}}:}}000^{0}101^{0}202^{0}\circ\circ\circn0n^{0}010^{1}111^{1}212^{1}\circ\circ\circn1n^{1}\circ\circ\circ0m0^{m}1m1^{m}2m2^{m}\circ\circ\circnmn^{m}aabbaabbaabbbbaaaabbaabbΣ\Sigmabbbbbbbbbbbbbbaaaabbaabb bad:{{{\mathpzc{B}}}}_{bad}:λ{\lambda}0{0}1{1}2{2}0011220,10,10,20,20,1,20,1,2𝒞bad:{{{\mathpzc{C}}}}_{bad}:λ{\lambda}0{0}1{1}2{2}3{3}0011220,1,30,1,30,20,20,1,2,30,1,2,3330,1,20,1,233𝒟bad:{{{\mathpzc{D}}}}_{bad}:λ{\lambda}0{0}1{1}2{2}3{3}4{4}0011220,1,30,1,30,2,40,2,40,1,2,3,40,1,2,3,43,43,40,1,2,30,1,2,33344440,1,2,3,40,1,2,3,444

Figure 4: On the left, a representative example for the Wagner class 𝔻𝕄n,m+\mathbb{DM}_{n,m}^{+}. The acceptance condition is ={{0i,1i,,ji}|j is odd iff i is odd}{{{\mathpzc{F}}}}=\{\{0^{i},1^{i},\ldots,j^{i}\}~|~j\mbox{ is odd iff }i\mbox{ is odd}\}. On the right a dba bad{{{\mathpzc{B}}}}_{bad} and dca 𝒞bad{{{\mathpzc{C}}}}_{bad} in 𝕀𝔹𝔹\mathbb{IB}\setminus\mathbb{RB} and 𝕀\mathbb{IC}\setminus\mathbb{RC}, respectively, and a dba 𝒟bad{{{\mathpzc{D}}}}_{bad} in 𝕀𝔹𝕀\mathbb{IB}\cap\mathbb{IC} that is not respective of its right congruence.
Proof 4.17.

Consider the dma 𝒟n,m+{{{\mathpzc{D}}}}_{n,m}^{+} in Figure 4, with acceptance condition ={{0i,1i,,ji}|j{{{\mathpzc{F}}}}=\{\{0^{i},1^{i},\ldots,j^{i}\}~|~j is odd iff ii is odd }\}. For instance, {00}\{0^{0}\}\in{{{\mathpzc{F}}}}, {00,10,20}\{0^{0},1^{0},2^{0}\}\in{{{\mathpzc{F}}}}, and {03,13}\{0^{3},1^{3}\}\in{{{\mathpzc{F}}}} but {10}\{1^{0}\}\notin{{{\mathpzc{F}}}} and {00,10}\{0^{0},1^{0}\}\notin{{{\mathpzc{F}}}}. It strictly belongs to the Wagner hierarchy class 𝔻𝕄n,m+\mathbb{DM}_{n,m}^{+}. To see that it is in 𝕀𝕄\mathbb{IM}, we show for each state, a word that distinguishes it from other states. We fix an order between the states: a state kk^{\ell} is smaller than kk^{\prime\ell^{\prime}} if either <\ell<\ell^{\prime} or =\ell=\ell^{\prime} and k<kk<k^{\prime}. Thus the last state nmn^{m} is the biggest in this order. The word aωa^{\omega} distinguishes the last state nmn^{m} from all states on odd rows, and the word (ab)ω(ab)^{\omega} distinguishes mnm^{n} from all states on even rows. For k[0..n]k\in[0..n], the word bnkaωb^{n-k}a^{\omega} distinguishes state kmk^{m} from all smaller states on odd rows, and the word bnk(ab)ωb^{n-k}(ab)^{\omega} distinguishes it from all smaller states on even rows. Finally, for k[0..n]k\in[0..n] and [0..m]\ell\in[0..m] the word (bn+1)mbnkaω(b^{n+1})^{m-\ell}b^{n-k}a^{\omega} distinguishes state kk^{\ell} from all smaller states on odd rows, and the word (bn+1)mbnk(ab)ω(b^{n+1})^{m-\ell}b^{n-k}(ab)^{\omega} distinguishes it from all smaller states on even rows. This shows that the rightcon automaton has (n+1)(m+1)(n+1)(m+1) states, and thus 𝒟n,m+𝕀𝕄{{\llbracket{{{\mathpzc{D}}}}_{n,m}^{+}\rrbracket}}\in\mathbb{IM}. It is also in 𝕀\mathbb{IP}, since we can define a dpa on the same structure, by assigning state kk^{\ell} the color kk if \ell is odd, and k+1k+1 if \ell is even. It is in 𝕀𝕋\mathbb{IT} since 𝕀𝕄𝕀𝕋\mathbb{IM}\subset\mathbb{IT}.

The proof for 𝔻𝕄n,m\mathbb{DM}_{n,m}^{-} is symmetric, and the proof for 𝔻𝕄n,m±\mathbb{DM}_{n,m}^{\pm} can be easily deduced from this.

5 Respective of the right congruence

As mentioned above, one of the motivations for studying classes of languages that are isomorphic to the right congruence is in the context of learning an unknown language. In this context, positive and negative examples (ω\omega-words labeled by their membership in the language) should help a learning algorithm to infer an automaton for the language. Consider the positive example (ab)ω(ab)^{\omega} for an unknown language LL. Intuitively, we expect that a minimal automaton for LL would have a loop of size 22 in which the word abab cycles. This is not necessarily the case, as shown by the language L=(aba+bab)ωL=(aba+bab)^{\omega}, whose minimal dba {{{\mathpzc{B}}}}_{\bowtie} is given in Fig 6. In the case of regular languages of finite words, if we regard the automaton {{{\mathpzc{B}}}}_{\bowtie} as a dfa, we note that abab, abababab are negative examples, while abababababab is a positive example. From this a learning algorithm can clearly infer the smallest loop on which abab cycles is of length 66, and not 22. But in the case of ω\omega-languages there are no negative ω\omega-words that can provide such information. We thus define a class of languages in which if uvωuv^{\omega} is a positive example for LL, then a minimal automaton for LL has a cycle of length at most |v||v| in which uvωuv^{\omega} loops.

Definition 5.18 (respective of L\sim_{L}).

A language LL is said to be respective of its right congruence if n0.\exists n_{0}\in\mathbb{N}. n>n0.\forall n>n_{0}. x,uΣ.\forall x,u\in\Sigma^{*}. xuωLxu^{\omega}\in L implies xunLxun+1xu^{n}\sim_{L}xu^{n+1}.

Intuitively, a language that is respective of its right congruence, can “delay” entering a loop as much as needed, but once it loops on a periodic part, it loops on the smallest period possible.

Being respective of the right congruence does not entail having an ω\omega-automaton that is isomorphic to the right congruence. Any language LL with |L|=1{|\sim_{L}|}=1 is (trivially) respective of its right congruence. By Proposition 3.5, L=(a+b)(aba)ωL=(a+b)^{*}(aba)^{\omega} has |L|=1{|\sim_{L}|}=1, but LL is not in 𝕀𝕋\mathbb{IT}, 𝕀𝕄\mathbb{IM}, 𝕀\mathbb{IP}, 𝕀𝔹\mathbb{IB} or 𝕀\mathbb{IC}, because every ω\omega-automaton accepting LL requires more than one state. We thus concentrate on languages which are both isomorphic to the rightcon automaton and respective of their right congruence. We use 𝔹\mathbb{RB}, \mathbb{RC}, \mathbb{RP}, 𝕄\mathbb{RM} and 𝕋\mathbb{RT} for the classes of languages that are respective of their right congruence and reside in 𝕀𝔹\mathbb{IB}, 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM} and 𝕀𝕋\mathbb{IT}, respectively. By definition, thus, 𝕀𝕏𝕏\mathbb{IX}\supseteq\mathbb{RX} for 𝕏{𝔹,,,𝕄,𝕋}\mathbb{X}\in\{\mathbb{B},\mathbb{C},\mathbb{P},\mathbb{M},\mathbb{T}\}. We show that these inclusions are strict.

Proposition 5.19.

𝕀𝔹𝔹\mathbb{IB}\supsetneq\mathbb{RB}, 𝕀\mathbb{IC}\supsetneq\mathbb{RC}, 𝕀\mathbb{IP}\supsetneq\mathbb{RP}, 𝕀𝕄𝕄\mathbb{IM}\supsetneq\mathbb{RM} and 𝕀𝕋𝕋\mathbb{IT}\supsetneq\mathbb{RT}.

Proof 5.20.

Consider the dba bad{{{\mathpzc{B}}}}_{bad} in Fig. 4 on the right. The language BbadB_{bad} accepted by bad{{{\mathpzc{B}}}}_{bad} is in 𝕀𝔹𝕀𝕀𝕄𝕀𝕋\mathbb{IB}\subset\mathbb{IP}\subset\mathbb{IM}\subset\mathbb{IT} but it is not respective of its right congruence. To see that it is in 𝕀𝔹\mathbb{IB} take ϵ\epsilon, 00, 11 and 22 as the representative words for states λ\lambda, 00, 11 and 22 respectively. Note that (011)ω(011)^{\omega} distinguishes 11 from the rest of the representative words, and (022)ω(022)^{\omega} distinguishes 22 from the rest of the representative words. Finally, (11)ω(11)^{\omega} distinguishes ϵ\epsilon from 00. The pair (ϵ,1012)(\epsilon,1012) shows that BbadB_{bad} is not respective of its right congruence since (1012)ωBbad(1012)^{\omega}\in B_{bad} yet for all nn\in\mathbb{N} we have that (1012)n+1≁Bbad(1012)n(1012)^{n+1}\not\sim_{B_{bad}}(1012)^{n}.

Consider the dca 𝒞bad{{{\mathpzc{C}}}}_{bad} in Fig. 4. The language CbadC_{bad} accepted by 𝒞bad{{{\mathpzc{C}}}}_{bad} is in 𝕀\mathbb{IC} but it is not respective of its right congruence. To see that it is in 𝕀\mathbb{IC} take ϵ\epsilon, 00, 11, 22 and 1313 as the representative words for states λ\lambda, 00, 11, 22 and 33 respectively. Note that 0ω0^{\omega} distinguishes 11 and 22 from the rest of the representative words; 3ω3^{\omega} distinguishes 11 from 22 and it distinguishes ϵ\epsilon from 00 and 1313; and 30ω30^{\omega} distinguishes 1313 from 00. The pair (ϵ,1012)(\epsilon,1012) shows that CbadC_{bad} is not respective of its right congruence since (1012)ωCbad(1012)^{\omega}\in C_{bad} yet for all nn\in\mathbb{N} we have that (1012)n+1≁Cbad(1012)n(1012)^{n+1}\not\sim_{C_{bad}}(1012)^{n}.

Recall that 𝔻𝔹𝔻=𝕀𝔹𝕀\mathbb{DB}\cap\mathbb{DC}=\mathbb{IB}\cap\mathbb{IC}. The dba 𝒟bad{{{\mathpzc{D}}}}_{bad} in Fig. 4 can be used to show a language in 𝕀𝔹𝕀\mathbb{IB}\cap\mathbb{IC} that is not respective of its right congruence.

Proposition 5.21.

There exists languages in 𝕀𝔹𝕀\mathbb{IB}\cap\mathbb{IC} that are not respective of their right congruences.

To complete the picture of inclusions between the 𝕏\mathbb{RX} classes, we establish that 𝕄\mathbb{RM}\supsetneq\mathbb{RP} and 𝕋𝕄\mathbb{RT}\supsetneq\mathbb{RM}.

Proposition 5.22.

𝕄\mathbb{RM}\supsetneq\mathbb{RP} and 𝕋𝕄\mathbb{RT}\supsetneq\mathbb{RM}.

Figure 1 on the right summarizes the above results. While it is notable that the requirement of being respective of the right congruence constitutes a restriction, there exist languages which are respective of their right congruence in every class of the Wagner Hierarchy.

Proposition 5.23.

Let n,mn,m be two natural numbers and p{+,,±}p\in\{+,-,\pm\}. Let 𝔻𝕄n,mp\mathbb{DM}_{n,m}^{p} denote the Wagner Hierarchy class with a maximum of nn alternations in each inclusion chain, and a sequence of at most mm chains of alternating polarity. Then there exists a language LΣωL\in\Sigma^{\omega} for Σ={a,b}\Sigma=\{a,b\} such that L𝔻𝕄n,mp𝕄𝕋L\in\mathbb{DM}_{n,m}^{p}\cap\mathbb{RM}\cap\mathbb{RP}\cap\mathbb{RT}.

Proof 5.24.

Consider again the dma 𝒟n,m+{{{\mathpzc{D}}}}_{n,m}^{+} in Fig. 4 and let Ln,m+{L_{n,m}^{+}} be the language that it recognizes. We have established in the proof of Proposition 4.16 that 𝒟n,m+𝕀𝕄𝕀𝕀𝕋{{\llbracket{{{\mathpzc{D}}}}_{n,m}^{+}\rrbracket}}\in\mathbb{IM}\cap\mathbb{IP}\cap\mathbb{IT}. Take n0=(m+1)(n+1)n_{0}=(m+1)(n+1). Consider the state that 𝒟n,m+{{{\mathpzc{D}}}}_{n,m}^{+} reaches after reading xun0xu^{n_{0}}. If this state is nmn^{m} then clearly for any n>n0n^{\prime}>n_{0}, xunxu^{n^{\prime}} also reaches states nmn^{m}. Otherwise there is an aa following the longest subsequence of bb^{*} in uu. The state that 𝒟n,m+{{{\mathpzc{D}}}}_{n,m}^{+} reaches after reading xun0xu^{n_{0}} depends on (a) the the longest subsequence of bb’s in xx (b) the longest subsequence of bb’s in uu and (c) the number of consecutive bb’s in the rightmost subsequence of bb’s in uu. Since the parameter (a) depends only on xx and the parameters (b) and (c) remain the same in uiu^{i} for any ii\in\mathbb{N} we have that xun0+iLn,m+xun0+i+1xu^{n_{0}+i}\sim_{L_{n,m}^{+}}xu^{n_{0}+i+1}. Thus, the accepted language is respective of its right congruence.

Relation to Non-Counting Languages

The definition of respective of L\sim_{L} is reminiscent of the definition of non-counting languages [14]. A language LΣωL\subseteq\Sigma^{\omega} is said to be non-counting iff n0.n>n0.u,vΣ,wΣω.uvnwLuvn+1wL\exists n_{0}\in\mathbb{N}.\ \forall n>n_{0}.\ \forall u,v\in\Sigma^{*},\ w\in\Sigma^{\omega}.uv^{n}w\in L\iff uv^{n+1}w\in L.

Proposition 5.25.

If LL is non-counting then LL is respective of its right-congruence.

The converse does not hold. The set (aa)bω(aa)^{*}b^{\omega} is respective of L\sim_{L} but is not non-counting.

Proposition 5.26.

There exist languages that are respective of L\sim_{L} but are not non-counting.

One of the most commonly used temporal logics is Linear temporal logic (LTL) [28]. LTL formulas are non-counting [19, 35, 14, 29]. But, there are LTL formulas that characterize languages that are not in 𝕀𝕋\mathbb{IT} (and thus not in any of the 𝕀\mathbb{I} classes). Indeed, the formula FG(aXa)FG\,(a\vee X\,a) characterizes the language L=Σ(a+Σa)ωL=\Sigma^{*}(a+\Sigma a)^{\omega} and by Proposition 3.5, |L|=1{|\sim_{L}|=1}.

6 Closure properties

We examine what Boolean closure properties hold or do not hold for these classes of languages recognizable by an automaton isomorphic to the rightcon automaton, and by automata that are also respective of the right congruence.

It is a well known result that weak regular ω\omega-languages are closed under all Boolean operations as stated by the following proposition.

Proposition 6.27 (c.f. [25]).

The class 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} is closed under the Boolean operations complementation, union and intersection.

The classes 𝕀𝕋\mathbb{IT}, 𝕀𝕄\mathbb{IM} and 𝕀\mathbb{IP} are closed under complementation. The other classes that we consider are not closed under complementation. See Fig. 5 for counterexamples.

Proposition 6.28.

The classes 𝕀𝕋\mathbb{IT}, 𝕀𝕄\mathbb{IM} and 𝕀\mathbb{IP} are closed under complementation.
The classes 𝕀𝔹\mathbb{IB} and 𝕀\mathbb{IC} are not closed under complementation.

Proposition 6.29.

𝔹\mathbb{RB}, \mathbb{RC}, \mathbb{RP}, 𝕄\mathbb{RM} and 𝕋\mathbb{RT} are not closed under complementation.

Proof 6.30.

Consider the dba 𝒫{{{\mathpzc{P}}}} in Fig. 5. The ω\omega-words ba(ac)ωba(ac)^{\omega}, a(ac)ωa(ac)^{\omega}, (ac)ω(ac)^{\omega} and (ca)ω(ca)^{\omega} distinguish its five states (the four shown in the figure, and the sink state). This shows that 𝒫𝕀𝔹𝕀𝕀𝕄𝕀𝕋{{{\mathpzc{P}}}}\in\mathbb{IB}\subset\mathbb{IP}\subset\mathbb{IM}\subset\mathbb{IT}. To see that it is respective of its right congruence, if xyωxy^{\omega} is in 𝒫{{\llbracket{{{\mathpzc{P}}}}\rrbracket}} then yy must traverse the cycle of states 3 and 4. Thus for some n0n_{0}, 𝒫(xyn0)=3{{{\mathpzc{P}}}}(xy^{n_{0}})=3 and y=(ac)ky=(ac)^{k} for some k1k\geq 1, or 𝒫(xyn0)=4{{{\mathpzc{P}}}}(xy^{n_{0}})=4 and y=(ca)ky=(ca)^{k} for some k1k\geq 1. In either case, 𝒫(xyn)=𝒫(xyn+1){{{\mathpzc{P}}}}(xy^{n})={{{\mathpzc{P}}}}(xy^{n+1}) for all nn0n\geq n_{0}. However, its complement 𝒫c{{{\mathpzc{P}}}}^{c} accepts the word bωb^{\omega}. But for every nn\in\mathbb{N} we have bn≁𝒫cbn+1b^{n}\not\sim_{{{\llbracket{{{\mathpzc{P}}}}^{c}\rrbracket}}}b^{n+1}.

This shows that 𝔹\mathbb{RB}, \mathbb{RP}, 𝕄\mathbb{RM} and 𝕋\mathbb{RT} are not closed under complementation. Consider the dca 𝒫{{{\mathpzc{P}}}}^{\prime} obtained from 𝒫{{{\mathpzc{P}}}} by marking the states {1,2,0}\{1,2,0\} where 00 is the sink state. Then 𝒫{{{\mathpzc{P}}}}^{\prime} recognizes the same language as 𝒫{{{\mathpzc{P}}}}. We get that 𝒫{{{\mathpzc{P}}}}^{\prime} is in 𝕀\mathbb{IC}, is respective of its right congruence, yet its complement is not respective of its right congruence.

Aside from the class 𝔻𝔹𝔻\mathbb{DB}\cap\mathbb{DC} none of the classes are closed under union or intersection. The automata in Figures 5 and 6 are used to refute these closures. The complete proofs are given in the full version.

Proposition 6.31.

The classes 𝕀𝔹\mathbb{IB}, 𝕀\mathbb{IC}, 𝕀\mathbb{IP}, 𝕀𝕄\mathbb{IM} and 𝕀𝕋\mathbb{IT} are not closed under union or intersection. The classes 𝔹\mathbb{RB}, \mathbb{RC}, \mathbb{RP}, 𝕄\mathbb{RM} and 𝕋\mathbb{RT} are not closed under union or intersection.

1:{{{\mathpzc{B}}}}_{1}:1{1}2{2}3{3}aabbaabbaabb2:{{{\mathpzc{B}}}}_{2}:1{1}2{2}3{3}bbaabbaaaabb/𝒞:{{{\mathpzc{B}}}}/{{{\mathpzc{C}}}}:001122aabbaabb𝒫:{{{\mathpzc{P}}}}:11223344a,ca,cbbb,cb,caaaacc

Figure 5: On the left examples for non-closure of union for 𝕀𝔹\mathbb{IB}. On the right a dba {{{\mathpzc{B}}}} and a dca 𝒞{{{\mathpzc{C}}}} showing 𝕀𝔹\mathbb{IB} and 𝕀\mathbb{IC} are not closed under complementation (the sink state is not shown).

1:{{0}}{{\mathpzc{M}}}_{1}:\{\{0\}\}𝒫1:κ(0)=1,κ(1)=2{{{\mathpzc{P}}}}_{1}:\kappa(0)=1,\kappa(1)=2𝒞1:{1}{{{\mathpzc{C}}}}_{1}:\{1\}0{0}1{1}b,cb,caaa,ca,cbb2:{{1}}{{\mathpzc{M}}}_{2}:\{\{1\}\}𝒫2:κ(0)=2,κ(1)=1{{{\mathpzc{P}}}}_{2}:\kappa(0)=2,\kappa(1)=1𝒞2:{0}{{{\mathpzc{C}}}}_{2}:\{0\}0{0}1{1}b,cb,caaa,ca,cbb3:{{0},{1}}{{\mathpzc{M}}}_{3}:\{\{0\},\{1\}\}0{0}1{1}b,cb,caaa,ca,cbb:{{{\mathpzc{B}}}}_{\bowtie}:0{0}1{1}2{2}3{3}4{4}aabbaabbaabb

Figure 6: On the left, examples for non closure of union for 𝕀𝕄\mathbb{IM}, 𝕀\mathbb{IP}, and 𝕀\mathbb{IC}. On the right {{{\mathpzc{B}}}}_{\bowtie}.

7 Discussion

We have explored properties of the right congruences of regular ω\omega-languages, characterized when a language has a trivial right congruence, defined classes of languages that have a fully informative right congruence, and defined an orthogonal property of a language being respective of its right congruence, which is implied by but does not imply the property of being non-counting. We have shown that there are languages with fully informative right congruences in every class of the infinite Wagner hierarchy, and that this remains true if we consider languages that are also respective of their right congruences. The (mostly) non-closure results under Boolean operations are not necessarily inimical to learnability. Our hope is that future research will be able to take advantage of these properties in the search for efficient minimization and learning algorithms for regular ω\omega-languages.

References

  • [1]
  • [2] F. Aarts & F.W. Vaandrager (2010): Learning I/O Automata. In: CONCUR 2010 - Concurrency Theory, 21th Inter. Conf., CONCUR 2010, Paris, France, August 31-September 3, 2010. Proc., pp. 71–85, 10.1007/978-3-642-15375-4_6.
  • [3] D. Angluin (1987): Learning Regular Sets from Queries and Counterexamples. Inf. Comput. 75(2), pp. 87–106, 10.1016/0890-5401(87)90052-6.
  • [4] D. Angluin, U. Boker & D. Fisman (2016): Families of DFAs as Acceptors of omega-Regular Languages. In: 41st Inter. Symp. on Mathematical Foundations of Computer Science, MFCS, pp. 11:1–11:14, 10.4230/LIPIcs.MFCS.2016.11.
  • [5] D. Angluin, U. Boker & D. Fisman (2018): Families of DFAs as Acceptors of ω\omega-Regular Languages. Logical Methods in Computer Science Volume 14, Issue 1, 10.23638/LMCS-14(1:15)2018.
  • [6] D. Angluin & D. Fisman (2014): Learning Regular Omega Languages. In: Algorithmic Learning Theory - 25th Inter. Conf., ALT Proc., pp. 125–139, 10.1007/978-3-319-11662-4_10.
  • [7] D. Angluin & D. Fisman (2016): Learning regular omega languages. Theor. Comput. Sci. 650, pp. 57–72, 10.1016/j.tcs.2016.07.031.
  • [8] B. Balle & M. Mohri (2015): Learning Weighted Automata. In: Algebraic Informatics - 6th International Conference, CAI 2015, pp. 1–21, 10.1007/978-3-319-23021-4_1.
  • [9] F. Bergadano & S. Varricchio (1996): Learning Behaviors of Automata from Multiplicity and Equivalence Queries. SIAM J. Comput. 25(6), pp. 1268–1280, 10.1137/S009753979326091X.
  • [10] A. W. Biermann & J. A. Feldman (1972): On the Synthesis of Finite-State Machines from Samples of Their Behavior. IEEE Trans. Comput. 21(6), pp. 592–597, 10.1109/TC.1972.5009015.
  • [11] M. Bojańczyk (2014): Transducers with Origin Information. In: Automata, Languages, and Programming - 41st Inter. Colloq., ICALP, pp. 26–37, 10.1007/978-3-662-43951-7_3.
  • [12] M. Botincan & D. Babic (2013): Sigma*: symbolic learning of input-output specifications. In: 40th ACM SIGPLAN-SIGACT Symp. on Principles of Programming Language (POPL13), pp. 443–456, 10.1145/2429069.2429123.
  • [13] H. Calbrix, M. Nivat & A. Podelski (1994): Ultimately Periodic Words of Rational w-Languages. In: Proc. of the 9th Inter. Conf. on Mathematical Foundations of Programming Semantics, Springer-Verlag, pp. 554–566, 10.1007/3-540-58027-1_27.
  • [14] V. Diekert & P. Gastin (2008): First-order definable languages. In: Logic and Automata: History and Perspectives [in Honor of Wolfgang Thomas]., pp. 261–306.
  • [15] D. Drews & L. D’Antoni (2017): Learning Symbolic Automata. In: Tools and Algorithms for the Construction and Analysis of Systems - 23rd Inter. Conf., TACAS, pp. 173–189, 10.1007/978-3-662-54577-5_10.
  • [16] A. Farzan, Y. Chenand E.M. Clarke, Y. Tsay & B. Wang (2008): Extending Automated Compositional Verification to the Full Class of Omega-Regular Languages. In: Tools and Algorithms for the Construction and Analysis of Systems, LNCS 4963, Springer Berlin Heidelberg, pp. 2–17, 10.1007/978-3-540-78800-3_2.
  • [17] D. Fisman (2018): Inferring Regular Languages and ω\omega-Languages. Journal of Logical and Algebraic Methods in Programming 98C, pp. 27–49, 10.1016/j.jlamp.2018.03.002.
  • [18] M. Fliess (1974): Matrices de Hankel. Journal de Matheḿatiques Pureset Appliqueés, pp. 197–224.
  • [19] D. M. Gabbay, A. Pnueli, S. Shelah & J. Stavi (1980): On the Temporal Basis of Fairness. In: Conf. Record of the 7th Annual ACM Symp. on Principles of Programming Languages, pp. 163–173, 10.1145/567446.567462.
  • [20] O. Kupferman, G. Morgenstern & A. Murano (2006): Typeness for omega-regular Automata. Int. J. Found. Comput. Sci. 17(4), pp. 869–884, 10.1142/S0129054106004157.
  • [21] M. Leucker (2006): Learning Meets Verification. In: Formal Methods for Components and Objects, 5th Inter. Symp., FMCO, pp. 127–151, 10.1007/978-3-540-74792-5_6.
  • [22] O. Maler & I. Mens (2017): A Generic Algorithm for Learning Symbolic Automata from Membership Queries. In: Models, Algorithms, Logics and Tools - Essays Dedicated to Kim Guldstrand Larsen on the Occasion of His 60th Birthday, pp. 146–169, 10.1007/978-3-319-63121-9_8.
  • [23] O. Maler & A. Pnueli (1995): On the Learnability of Infinitary Regular Sets. Inf. Comput. 118(2), pp. 316–326, 10.1006/inco.1995.1070.
  • [24] O. Maler & L. Staiger (1997): On Syntactic Congruences for Omega-Languages. Theor. Comput. Sci. 183(1), pp. 93–112, 10.1016/S0304-3975(96)00312-X.
  • [25] Z. Manna & A. Pnueli (1990): A Hierarchy of Temporal Properties. In: Proc. of the Ninth Annual ACM Symp. on Principles of Distributed Computing, pp. 377–410, 10.1145/93385.93442.
  • [26] I. Mens & O. Maler (2015): Learning Regular Languages over Large Ordered Alphabets. Logical Methods in Computer Science 11(3), 10.2168/LMCS-11(3:13)2015.
  • [27] D. Perrin & J-E. Pin (2004): Infinite Words: Automata, Semigroups, Logic and Games. Pure and Applied Math 141, Elsevier.
  • [28] A. Pnueli (1977): The Temporal Logic of Programs. In: FOCS, pp. 46–57, 10.1109/SFCS.1977.32.
  • [29] A. Rabinovich (2014): A Proof of Kamp’s theorem. Logical Methods in Computer Science 10(1), 10.2168/LMCS-10(1:14)2014.
  • [30] B. Le Saëc (1990): Saturating right congruences. ITA 24, pp. 545–560.
  • [31] S. Schewe (2010): Beyond Hyper-Minimisation—Minimising DBAs and DPAs is NP-Complete. In: IARCS Annual Conf. on Foundations of Software Technology and Theor. Comp. Science, FSTTCS, pp. 400–411, 10.4230/LIPIcs.FSTTCS.2010.400.
  • [32] L. Staiger (1983): Finite-State omega-Languages. J. Comput. Syst. Sci. 27(3), pp. 434–448, 10.1016/0022-0000(83)90051-X.
  • [33] D. Long Van, B. Le Saëc & I. Litovsky (1995): Characterizations of Rational omega-Languages by Means of Right Congruences. Theor. Comput. Sci. 143(1), pp. 1–21, 10.1016/0304-3975(95)80022-2.
  • [34] K. W. Wagner (1975): A Hierarchy of Regular Sequence Sets. In: MFCS, pp. 445–449, 10.1007/3-540-07389-2_231.
  • [35] P. Wolper (1983): Temporal Logic Can Be More Expressive. Information and Control 56(1/2), pp. 72–99, 10.1016/S0019-9958(83)80051-5.