1 |
«Con la grande polvareda» ; «Con la grande polvareda»the ballad of La pérdida de Don Beltrán in the spanish golden age ; el romance de La pérdida de don Beltrán en el Siglo de Oro
|
|
|
|
BASE
|
|
Show details
|
|
2 |
Formal Language Recognition by Hard Attention Transformers: Perspectives from Circuit Complexity ...
|
|
|
|
BASE
|
|
Show details
|
|
3 |
Towards a Semantic Information Theory (Introducing Quantum Corollas) ...
|
|
|
|
BASE
|
|
Show details
|
|
6 |
Elastic Full Procrustes Analysis of Plane Curves via Hermitian Covariance Smoothing ...
|
|
|
|
BASE
|
|
Show details
|
|
10 |
Grammatical Gender Disambiguates Syntactically Similar Nouns
|
|
|
|
In: Entropy; Volume 24; Issue 4; Pages: 520 (2022)
|
|
BASE
|
|
Show details
|
|
11 |
Optimal alphabet for single text compression ...
|
|
|
|
Abstract:
A text can be viewed via different representations, i.e. as a sequence of letters, n-grams of letters, syllables, words, and phrases. Here we study the optimal noiseless compression of texts using the Huffman code, where the alphabet of encoding coincides with one of those representations. We show that it is necessary to account for the codebook when compressing a single text. Hence, the total compression comprises of the optimally compressed text -- characterized by the entropy of the alphabet elements -- and the codebook which is text-specific and therefore has to be included for noiseless (de)compression. For texts of Project Gutenberg the best compression is provided by syllables, i.e. the minimal meaning-expressing element of the language. If only sufficiently short texts are retained, the optimal alphabet is that of letters or 2-grams of letters depending on the retained length. ... : 11 pages, 12 figures, 1 table ...
|
|
Keyword:
Computation and Language cs.CL; Data Analysis, Statistics and Probability physics.data-an; FOS Computer and information sciences; FOS Physical sciences; Information Theory cs.IT
|
|
URL: https://arxiv.org/abs/2201.05234 https://dx.doi.org/10.48550/arxiv.2201.05234
|
|
BASE
|
|
Hide details
|
|
12 |
Gegen die Öffentlichkeit: Alternative Nachrichtenmedien im deutschsprachigen Raum
|
|
Schwaiger, Lisa. - : transcript Verlag, 2022. : DEU, 2022. : Bielefeld, 2022
|
|
In: 46 ; Digitale Gesellschaft ; 327 (2022)
|
|
BASE
|
|
Show details
|
|
13 |
“What Can I Cook with These Ingredients?” - Understanding Cooking-Related Information Needs in Conversational Search
|
|
|
|
BASE
|
|
Show details
|
|
19 |
A Developmental Framework for Embodiment Research: The Next Step Toward Integrating Concepts and Methods.
|
|
|
|
BASE
|
|
Show details
|
|
20 |
A quantitative perspective on Japanese accent
|
|
|
|
In: 34th Paris Meeting on East Asian Linguistics ; https://hal.archives-ouvertes.fr/hal-03283679 ; 34th Paris Meeting on East Asian Linguistics, Jul 2021, Paris, France (2021)
|
|
BASE
|
|
Show details
|
|
|
|