Portuguese part-of-speech tagging with large margin structure learning

Research output: Contributions to collected editions/worksArticle in conference proceedingsResearchpeer-review

Standard

Portuguese part-of-speech tagging with large margin structure learning. / Fernandes, Eraldo Rezende; Rodrigues, Irving Muller; Milidiú, Ruy Luiz.
BRACIS 2014: 2014 Brazilian Conference on Intelligent Systems ; 19-23 October 2014, São Carlos, São Paulo, Brazil ; proceedings. Piscataway: Institute of Electrical and Electronics Engineers Inc., 2014. p. 25-30 6984802.

Research output: Contributions to collected editions/worksArticle in conference proceedingsResearchpeer-review

Harvard

Fernandes, ER, Rodrigues, IM & Milidiú, RL 2014, Portuguese part-of-speech tagging with large margin structure learning. in BRACIS 2014: 2014 Brazilian Conference on Intelligent Systems ; 19-23 October 2014, São Carlos, São Paulo, Brazil ; proceedings., 6984802, Institute of Electrical and Electronics Engineers Inc., Piscataway, pp. 25-30, Brazilian Conference on Intelligent Systems - BRACIS 2014, Sao Carlos, Sao Paulo, Brazil, 18.10.14. https://doi.org/10.1109/BRACIS.2014.16

APA

Fernandes, E. R., Rodrigues, I. M., & Milidiú, R. L. (2014). Portuguese part-of-speech tagging with large margin structure learning. In BRACIS 2014: 2014 Brazilian Conference on Intelligent Systems ; 19-23 October 2014, São Carlos, São Paulo, Brazil ; proceedings (pp. 25-30). Article 6984802 Institute of Electrical and Electronics Engineers Inc.. https://doi.org/10.1109/BRACIS.2014.16

Vancouver

Fernandes ER, Rodrigues IM, Milidiú RL. Portuguese part-of-speech tagging with large margin structure learning. In BRACIS 2014: 2014 Brazilian Conference on Intelligent Systems ; 19-23 October 2014, São Carlos, São Paulo, Brazil ; proceedings. Piscataway: Institute of Electrical and Electronics Engineers Inc. 2014. p. 25-30. 6984802 doi: 10.1109/BRACIS.2014.16

Bibtex

@inbook{8908bae32b724634be8419587a8cc4be,
title = "Portuguese part-of-speech tagging with large margin structure learning",
abstract = "Part-of-Speech Tagging is a fundamental task on many Natural Language Processing systems. This task consists in identifying the syntactic category, i.e. the part of speech, of each word in a sentence. Despite the fact that the current state-of-the-art accuracy for this task is around 97%, any improvement has an immediate impact on more complex tasks, like Parsing, Semantic Role Labeling and Information Extraction. Thus, it is still relevant to explore this task. In this paper, we introduce a part-of-speech tagger based on the Structure Learning framework that reduces the smallest known error on the Portuguese Mac-Morpho corpus by 7.8%. We also apply our tagger to a recently revised version of Mac-Morpho. Our system accuracy on this latter version is competitive with a semi-supervised Neural Network trained on Mac-Morpho plus a very large non-annotated corpus. Additionally, our system is simpler than previous systems and uses a very limited feature set. Our system employs a Large Margin training criteria to derive a structure predictor that is more robust on unseen data.",
keywords = "Machine Learning, Natural Language Processing, POS Tagging, Structure Learning, Informatics, Business informatics",
author = "Fernandes, {Eraldo Rezende} and Rodrigues, {Irving Muller} and Milidi{\'u}, {Ruy Luiz}",
year = "2014",
month = dec,
day = "12",
doi = "10.1109/BRACIS.2014.16",
language = "English",
isbn = "978-1-4799-7859-5",
pages = "25--30",
booktitle = "BRACIS 2014",
publisher = "Institute of Electrical and Electronics Engineers Inc.",
address = "United States",
note = "Brazilian Conference on Intelligent Systems - BRACIS 2014 ; Conference date: 18-10-2014 Through 23-10-2014",
url = "https://ieeexplore.ieee.org/xpl/conhome/6979382/proceeding",

}

RIS

TY - CHAP

T1 - Portuguese part-of-speech tagging with large margin structure learning

AU - Fernandes, Eraldo Rezende

AU - Rodrigues, Irving Muller

AU - Milidiú, Ruy Luiz

N1 - Conference code: 3

PY - 2014/12/12

Y1 - 2014/12/12

N2 - Part-of-Speech Tagging is a fundamental task on many Natural Language Processing systems. This task consists in identifying the syntactic category, i.e. the part of speech, of each word in a sentence. Despite the fact that the current state-of-the-art accuracy for this task is around 97%, any improvement has an immediate impact on more complex tasks, like Parsing, Semantic Role Labeling and Information Extraction. Thus, it is still relevant to explore this task. In this paper, we introduce a part-of-speech tagger based on the Structure Learning framework that reduces the smallest known error on the Portuguese Mac-Morpho corpus by 7.8%. We also apply our tagger to a recently revised version of Mac-Morpho. Our system accuracy on this latter version is competitive with a semi-supervised Neural Network trained on Mac-Morpho plus a very large non-annotated corpus. Additionally, our system is simpler than previous systems and uses a very limited feature set. Our system employs a Large Margin training criteria to derive a structure predictor that is more robust on unseen data.

AB - Part-of-Speech Tagging is a fundamental task on many Natural Language Processing systems. This task consists in identifying the syntactic category, i.e. the part of speech, of each word in a sentence. Despite the fact that the current state-of-the-art accuracy for this task is around 97%, any improvement has an immediate impact on more complex tasks, like Parsing, Semantic Role Labeling and Information Extraction. Thus, it is still relevant to explore this task. In this paper, we introduce a part-of-speech tagger based on the Structure Learning framework that reduces the smallest known error on the Portuguese Mac-Morpho corpus by 7.8%. We also apply our tagger to a recently revised version of Mac-Morpho. Our system accuracy on this latter version is competitive with a semi-supervised Neural Network trained on Mac-Morpho plus a very large non-annotated corpus. Additionally, our system is simpler than previous systems and uses a very limited feature set. Our system employs a Large Margin training criteria to derive a structure predictor that is more robust on unseen data.

KW - Machine Learning

KW - Natural Language Processing

KW - POS Tagging

KW - Structure Learning

KW - Informatics

KW - Business informatics

UR - http://www.scopus.com/inward/record.url?scp=84922535000&partnerID=8YFLogxK

U2 - 10.1109/BRACIS.2014.16

DO - 10.1109/BRACIS.2014.16

M3 - Article in conference proceedings

AN - SCOPUS:84922535000

SN - 978-1-4799-7859-5

SP - 25

EP - 30

BT - BRACIS 2014

PB - Institute of Electrical and Electronics Engineers Inc.

CY - Piscataway

T2 - Brazilian Conference on Intelligent Systems - BRACIS 2014

Y2 - 18 October 2014 through 23 October 2014

ER -

DOI

Recently viewed

Publications

  1. Qualitative Daten computergestutzt auswerten
  2. Non-technical success factors for bioenergy projects-Learning from a multiple case study in Japan
  3. Making the matrix matter
  4. The Challenge of Democratic Representation in the European Union
  5. Geometric structures using model predictive control for an electromagnetic actuator
  6. Integrating regional perceptions into climate change adaptation
  7. Theory-based course design for professional master's degree program in business engineering
  8. Implementation of formative assessment
  9. Organizing Events for Configuring and Maintaining Creative Fields
  10. Arc spraying of WCFeCSiMn cored wires.
  11. Multiscale solutions of the electromagnetic continuity differential equation using packets of harmonic wavelets
  12. Operationalizing ecosystem services for the mitigation of soil threats
  13. Deconstructing and reconstructing diversity in client-provider-relationships of social work
  14. (Un)Bestimmtheit
  15. The Impact of Mental Fatigue on Exploration in a Complex Computer Task
  16. Script and sound
  17. Using Multi-Label Classification for Improved Question Answering
  18. Predicting online user behavior based on Real-Time Advertising Data
  19. Question Answering Mediated by Visual Clues and Knowledge Graphs
  20. From Learning Machines to Learning Humans
  21. Reducing problematic alcohol use in employees: economic evaluation of guided and unguided web-based interventions alongside a three-arm randomized controlled trial
  22. Logical-Rollenspiele
  23. Aboveground overyielding in grassland mixtures is associated with reduced biomass partitioning to belowground organs
  24. Green your community click by click
  25. "Introduction," communication +1
  26. On-board pneumatic pressure generation methods for soft robotics applications
  27. Degrees of Integration
  28. Does online-delivered Cognitive Behavioural Therapy for Insomnia improve insomnia severity in nurses working shifts? Protocol for a randomised-controlled trial