Yi-jie Tang


2014

pdf bib
FAdR: A System for Recognizing False Online Advertisements
Yi-jie Tang | Hsin-Hsi Chen
Proceedings of 52nd Annual Meeting of the Association for Computational Linguistics: System Demonstrations

pdf bib
Chinese Irony Corpus Construction and Ironic Structure Analysis
Yi-jie Tang | Hsin-Hsi Chen
Proceedings of COLING 2014, the 25th International Conference on Computational Linguistics: Technical Papers

2012

pdf bib
Advertising Legality Recognition
Yi-jie Tang | Cong-kai Lin | Hsin-Hsi Chen
Proceedings of COLING 2012: Posters

pdf bib
Mining Sentiment Words from Microblogs for Predicting Writer-Reader Emotion Transition
Yi-jie Tang | Hsin-Hsi Chen
Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC'12)

The conversations between posters and repliers in microblogs form a valuable writer-reader emotion corpus. This paper adopts a log relative frequency ratio to investigate the linguistic features which affect emotion transitions, and applies the results to predict writers' and readers' emotions. A 4-class emotion transition predictor, a 2-class writer emotion predictor, and a 2-class reader emotion predictor are proposed and compared.

pdf bib
Development of a Web-Scale Chinese Word N-gram Corpus with Parts of Speech Information
Chi-Hsin Yu | Yi-jie Tang | Hsin-Hsi Chen
Proceedings of the Eighth International Conference on Language Resources and Evaluation (LREC'12)

Web provides a large-scale corpus for researchers to study the language usages in real world. Developing a web-scale corpus needs not only a lot of computation resources, but also great efforts to handle the large variations in the web texts, such as character encoding in processing Chinese web texts. In this paper, we aim to develop a web-scale Chinese word N-gram corpus with parts of speech information called NTU PN-Gram corpus using the ClueWeb09 dataset. We focus on the character encoding and some Chinese-specific issues. The statistics about the dataset is reported. We will make the resulting corpus a public available resource to boost the Chinese language processing.

2011

pdf bib
Emotion Modeling from Writer/Reader Perspectives Using a Microblog Dataset
Yi-jie Tang | Hsin-Hsi Chen
Proceedings of the Workshop on Sentiment Analysis where AI meets Psychology (SAAIP 2011)