Registro completo de metadatos
| Campo DC | Valor | Lengua/Idioma |
|---|---|---|
| dc.rights.license | Reconocimiento 4.0 Internacional. (CC BY) | - |
| dc.contributor.author | da Silva, Juan | es |
| dc.contributor.author | Yovine, Sergio | es |
| dc.date.accessioned | 2026-08-17T18:45:52Z | - |
| dc.date.available | 2026-08-17T18:45:52Z | - |
| dc.date.issued | 2026-08 | - |
| dc.identifier.uri | https://hdl.handle.net/20.500.12381/5636 | - |
| dc.description.abstract | We extract visibly pushdown grammars (VPGs) from neural language models in a fully black-box setting. Our work builds on the VPL* framework, which learns VPGs from recurrent networks by exploiting access to the target's internal state. First, we replace the white-box equivalence oracle with PAC sampling over trees. The resulting algorithm applies unchanged to transformer language models and, more generally, to any acceptor exposing only binary decisions. Second, we further observe that the original framework only ever queries the target on well-formed sequences that can be accepted by the most permissive VPG (called BParse) with the target's ranked alphabet. That is, VPL* is blind to the target's behavior outside BParse. Thus, we propose to sample trees directly instead of sequences and call the target with all the resulting sequences in order to provide an error estimate of how far the target deviates from being a VPG. Last but not least, we develop a learner that captures a tree automaton describing discovered behaviors of the target outside BParse, which results in a larger learnable space. We evaluate the approach on several cases, including transformers trained with Dyck grammars and synthetic targets designed to be outside BParse. | es |
| dc.description.sponsorship | Agencia Nacional de Investigación e Innovación | es |
| dc.language.iso | eng | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3417 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3418 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3419 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3420 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3622 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3624 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3626 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/3656 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/5138 | es |
| dc.relation | https://doi.org/10.60895/redata/Z8QDEZ | es |
| dc.relation | https://doi.org/10.60895/redata/NDHQQQ | es |
| dc.relation | https://doi.org/10.60895/redata/KNERSJ | es |
| dc.relation | https://doi.org/10.60895/redata/JY5DUS | es |
| dc.relation | https://hdl.handle.net/20.500.12381/5632 | es |
| dc.relation | https://hdl.handle.net/20.500.12381/5633 | es |
| dc.rights | Acceso abierto | * |
| dc.subject | Active Learning | es |
| dc.subject | Tree Automata | es |
| dc.subject | Visibly Pushdown Languages | es |
| dc.title | Learning of Tree Automata applied to Neural Language Acceptors | es |
| dc.type | Preprint | es |
| dc.subject.anii | Ciencias Naturales y Exactas | |
| dc.subject.anii | Ciencias de la Computación e Información | |
| dc.identifier.anii | FMV_1_2023_1_175864 | es |
| dc.identifier.anii | POS_NAC_2023_1_178663 | es |
| dc.anii.institucionresponsable | Universidad ORT Uruguay | es |
| dc.anii.subjectcompleto | //Ciencias Naturales y Exactas/Ciencias de la Computación e Información/Ciencias de la Computación e Información | es |
| Aparece en las colecciones: | Publicaciones de ANII | |
Archivos en este ítem:
| archivo | Descripción | Tamaño | Formato | ||
|---|---|---|---|---|---|
| Learning_tree_automata__ICGI_2026___arXiv_.pdf | Descargar | 317.31 kB | Adobe PDF |
Las obras en REDI están protegidas por licencias Creative Commons.
Por más información sobre los términos de esta publicación, visita:
Reconocimiento 4.0 Internacional. (CC BY)
