Predicting CEFR levels in learners of English: The use of microsystem criterial features in a machine learning approach

ReCALL ◽

10.1017/s095834402100029x ◽

2021 ◽

pp. 1-17

Author(s):

Thomas Gaillat ◽

Andrew Simpkin ◽

Nicolas Ballier ◽

Bernardo Stearns ◽

Annanda Sousa ◽

...

Keyword(s):

Machine Learning ◽

Foreign Language ◽

Supervised Learning ◽

Language Proficiency ◽

Scoring System ◽

Learning Approach ◽

Linguistic Complexity ◽

Proficiency Levels ◽

External Data ◽

Machine Learning Approach

Abstract This paper focuses on automatically assessing language proficiency levels according to linguistic complexity in learner English. We implement a supervised learning approach as part of an automatic essay scoring system. The objective is to uncover Common European Framework of Reference for Languages (CEFR) criterial features in writings by learners of English as a foreign language. Our method relies on the concept of microsystems with features related to learner-specific linguistic systems in which several forms operate paradigmatically. Results on internal data show that different microsystems help classify writings from A1 to C2 levels (82% balanced accuracy). Overall results on external data show that a combination of lexical, syntactic, cohesive and accuracy features yields the most efficient classification across several corpora (59.2% balanced accuracy).

A Machine Learning Approach to detect Depression and Anxiety using Supervised Learning

2020 IEEE Asia-Pacific Conference on Computer Science and Data Engineering (CSDE) ◽

10.1109/csde50874.2020.9411642 ◽

2020 ◽

Author(s):

Anamika Ahmed ◽

Raihan Sultana ◽

Md Tahmidur Rahman Ullas ◽

Mariyam Begom ◽

Md. Muzahidul Islam Rahi ◽

...

Keyword(s):

Machine Learning ◽

Supervised Learning ◽

Learning Approach ◽

Depression And Anxiety ◽

Machine Learning Approach

Validation of a machine learning approach using FIB-4 and APRI scores assessed by the metavir scoring system: A cohort study

Arab Journal of Gastroenterology ◽

10.1016/j.ajg.2021.04.003 ◽

2021 ◽

Author(s):

Ahmed Hashem ◽

Abubakr Awad ◽

Hend Shousha ◽

Wafaa Alakel ◽

Ahmed Salama ◽

...

Keyword(s):

Machine Learning ◽

Cohort Study ◽

Scoring System ◽

Learning Approach ◽

System A ◽

Machine Learning Approach

Constructing and Validating Geographically Refined HAZUS-MH4 Hurricane Wind Risk Models: A Machine Learning Approach

Advances in Hurricane Engineering ◽

10.1061/9780784412626.092 ◽

2012 ◽

Cited By ~ 2

Author(s):

D. Subramanian ◽

J. Salazar ◽

L. Duenas-Osorio ◽

R. Stein

Keyword(s):

Machine Learning ◽

Learning Approach ◽

Risk Models ◽

Hurricane Wind ◽

Machine Learning Approach

The impact of economic plans on the Chinese education system: a machine learning approach

CADMO ◽

10.3280/cad2018-001005 ◽

2018 ◽

pp. 37-49

Author(s):

Wenjun Lin ◽

Xuefu Xu ◽

Francesco Dell’Anna

Keyword(s):

Machine Learning ◽

Education System ◽

Learning Approach ◽

Chinese Education ◽

System A ◽

Machine Learning Approach ◽

The Impact

Improving Bandwidth Utilization and Fairness between TCP Flows based on a Machine-learning Approach

IEEJ Transactions on Electronics Information and Systems ◽

10.1541/ieejeiss.133.1259 ◽

2013 ◽

Vol 133 (6) ◽

pp. 1259-1268

Author(s):

Akihiro Shiozu ◽

Syunji Yazaki ◽

K^|^ocirc;ki Abe

Keyword(s):

Machine Learning ◽

Learning Approach ◽

Bandwidth Utilization ◽

Machine Learning Approach

A Machine Learning Approach to Anaphora Resolution in Arabic

International Review on Computers and Software (IRECOS) ◽

10.15866/irecos.v9i12.4786 ◽

2014 ◽

Vol 9 (12) ◽

pp. 1956

Author(s):

Abdullatif Abolohom ◽

Nazlia Omar

Keyword(s):

Machine Learning ◽

Learning Approach ◽

Anaphora Resolution ◽

Machine Learning Approach

1552-P: Machine Learning Approach to Decision-Making for Initial Insulin Use in Japanese Patients with Type 2 Diabetes

Diabetes ◽

10.2337/db20-1552-p ◽

2020 ◽

Vol 69 (Supplement 1) ◽

pp. 1552-P

Author(s):

KAZUYA FUJIHARA ◽

MAYUKO H. YAMADA ◽

YASUHIRO MATSUBAYASHI ◽

MASAHIKO YAMAMOTO ◽

TOSHIHIRO IIZUKA ◽

...

Keyword(s):

Machine Learning ◽

Type 2 Diabetes ◽

Decision Making ◽

Japanese Patients ◽

Learning Approach ◽

Machine Learning Approach ◽

Insulin Use

A Machine Learning Approach to Jet-Surface Interaction Noise Modeling

AIAA Scitech 2020 Forum ◽

10.2514/6.2020-1728 ◽

2020 ◽

Author(s):

Clifford A. Brown ◽

Jonny Dowdall ◽

Brian Whiteaker ◽

Lauren McIntyre

Keyword(s):

Machine Learning ◽

Surface Interaction ◽

Learning Approach ◽

Noise Modeling ◽

Machine Learning Approach

A machine learning approach to evaluate IgE and IgG4 responses as patient-specific measure of exposure to sublingual allergen immunotherapy

10.26226/morressier.5afda3c8d64f25002cfc4098 ◽

2018 ◽

Author(s):

Thomas Stranzl

Keyword(s):

Machine Learning ◽

Allergen Immunotherapy ◽

Patient Specific ◽

Learning Approach ◽

Specific Measure ◽

Machine Learning Approach

Mol2vec: Unsupervised Machine Learning Approach with Chemical Intuition

10.26434/chemrxiv.5513581.v1 ◽

2017 ◽

Author(s):

Sabrina Jaeger ◽

Simone Fulle ◽

Samo Turk

Keyword(s):

Machine Learning ◽

Language Processing ◽

Supervised Machine Learning ◽

Learning Approach ◽

Learning Approaches ◽

Unsupervised Machine Learning ◽

Feature Representations ◽

Machine Learning Approach ◽

The Individual ◽

Vector Representations

Inspired by natural language processing techniques we here introduce Mol2vec which is an unsupervised machine learning approach to learn vector representations of molecular substructures. Similarly, to the Word2vec models where vectors of closely related words are in close proximity in the vector space, Mol2vec learns vector representations of molecular substructures that are pointing in similar directions for chemically related substructures. Compounds can finally be encoded as vectors by summing up vectors of the individual substructures and, for instance, feed into supervised machine learning approaches to predict compound properties. The underlying substructure vector embeddings are obtained by training an unsupervised machine learning approach on a so-called corpus of compounds that consists of all available chemical matter. The resulting Mol2vec model is pre-trained once, yields dense vector representations and overcomes drawbacks of common compound feature representations such as sparseness and bit collisions. The prediction capabilities are demonstrated on several compound property and bioactivity data sets and compared with results obtained for Morgan fingerprints as reference compound representation. Mol2vec can be easily combined with ProtVec, which employs the same Word2vec concept on protein sequences, resulting in a proteochemometric approach that is alignment independent and can be thus also easily used for proteins with low sequence similarities.