Evolving Text Data Stream Mining
–arXiv.org Artificial Intelligence
A text stream is an ordered sequence of text documents generated over time. A massive amount of such text data is generated by online social platforms every day. Designing an algorithm for such text streams to extract useful information is a challenging task due to unique properties of the stream such as infinite length, data sparsity, and evolution. Thereby, learning useful information from such streaming data under the constraint of limited time and memory has gained increasing attention. During the past decade, although many text stream mining algorithms have proposed, there still exists some potential issues. First, high-dimensional text data heavily degrades the learning performance until the model either works on subspace or reduces the global feature space. The second issue is to extract semantic text representation of documents and capture evolving topics over time. Moreover, the problem of label scarcity exists, whereas existing approaches work on the full availability of labeled data. To deal with these issues, in this thesis, new learning models are proposed for clustering and multi-label learning on text streams.
arXiv.org Artificial Intelligence
Aug-15-2024
- Country:
- South America > Brazil
- Maranhão (0.04)
- Oceania > Australia
- North America
- United States
- Minnesota > Hennepin County
- Minneapolis (0.13)
- Indiana > Marion County
- Indianapolis (0.04)
- Massachusetts > Suffolk County
- Boston (0.04)
- Hawaii > Honolulu County
- Honolulu (0.04)
- Virginia > Arlington County
- Arlington (0.04)
- Oregon
- Multnomah County > Portland (0.04)
- Benton County > Corvallis (0.04)
- Illinois > Cook County
- Chicago (0.04)
- Georgia > Fulton County
- Atlanta (0.04)
- Florida > Palm Beach County
- Boca Raton (0.04)
- California
- San Francisco County > San Francisco (0.14)
- Alameda County > Oakland (0.04)
- New York > New York County
- New York City (0.04)
- Minnesota > Hennepin County
- Canada
- Quebec > Montreal (0.04)
- Ontario > Toronto (0.04)
- Nova Scotia > Halifax Regional Municipality
- Halifax (0.04)
- British Columbia > Metro Vancouver Regional District
- Vancouver (0.04)
- United States
- Europe
- Germany > Berlin (0.04)
- Middle East > Cyprus (0.04)
- Czechia > Prague (0.04)
- Portugal > Porto
- Porto (0.04)
- Italy > Tuscany
- Florence (0.04)
- Pisa Province > Pisa (0.04)
- United Kingdom > England
- Greater London > London (0.04)
- Cambridgeshire > Cambridge (0.04)
- Bristol (0.04)
- Spain > Catalonia
- Barcelona Province > Barcelona (0.14)
- Norway > Central Norway
- Finland > Uusimaa
- Helsinki (0.04)
- Russia > Northwestern Federal District
- Leningrad Oblast > Saint Petersburg (0.04)
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Asia
- South Korea (0.04)
- Singapore (0.04)
- Russia (0.04)
- Middle East
- Japan > Honshū
- Kansai > Kyoto Prefecture
- Kyoto (0.04)
- Chūbu > Ishikawa Prefecture
- Kanazawa (0.04)
- Kansai > Kyoto Prefecture
- India > Telangana
- Hyderabad (0.04)
- China
- Yunnan Province > Kunming (0.04)
- Hong Kong (0.04)
- Beijing > Beijing (0.04)
- Heilongjiang Province > Harbin (0.04)
- South America > Brazil
- Genre:
- Overview (1.00)
- Research Report
- New Finding (1.00)
- Experimental Study (0.92)
- Promising Solution (0.67)
- Industry:
- Information Technology (1.00)
- Health & Medicine (1.00)
- Technology: