abstract class BaseUTF8Encoding extends UnicodeEncoding
| Modifier and Type | Field and Description |
|---|---|
private static int |
INVALID_CODE_FE |
private static int |
INVALID_CODE_FF |
(package private) static boolean |
USE_INVALID_CODE_SCHEME |
EMPTY_FOLD_CODESCHAR_INVALID, charset, hashCode, isAsciiCompatible, isDummy, isFixedWidth, isSingleByte, maxLength, minLength, name, NEW_LINE| Modifier | Constructor and Description |
|---|---|
protected |
BaseUTF8Encoding(int[] EncLen,
int[][] Trans) |
| Modifier and Type | Method and Description |
|---|---|
int |
codeToMbc(int code,
byte[] bytes,
int p)
Extracts code point into it's multibyte representation
|
int |
codeToMbcLength(int code)
Returns character length given a code point
Oniguruma equivalent:
code_to_mbclen |
int[] |
ctypeCodeRange(int ctype,
IntHolder sbOut)
utf8_get_ctype_code_range
|
java.lang.String |
getCharsetName() |
boolean |
isNewLine(byte[] bytes,
int p,
int end)
onigenc_is_mbc_newline_0x0a / used also by multibyte encodings
|
boolean |
isReverseMatchAllowed(byte[] bytes,
int p,
int end)
onigenc_always_true_is_allowed_reverse_match
|
int |
leftAdjustCharHead(byte[] bytes,
int p,
int s,
int end)
utf8_left_adjust_char_head
|
int |
mbcCaseFold(int flag,
byte[] bytes,
IntHolder pp,
int end,
byte[] fold)
onigenc_ascii_mbc_case_fold
|
int |
mbcToCode(byte[] bytes,
int p,
int end)
Returns code point for a character
Oniguruma equivalent:
mbc_to_code |
(package private) static byte |
trail0(int code) |
(package private) static byte |
trailS(int code,
int shift) |
private static boolean |
utf8IsLead(int c) |
applyAllCaseFold, caseFoldCodesByString, ctypeCodeRange, isCodeCType, propertyNameToCTypelength, lengthForTwoUptoFour, mb2CodeToMbc, mb2CodeToMbcLength, mb2IsCodeCType, mb4CodeToMbc, mb4CodeToMbcLength, mb4IsCodeCType, mbnMbcCaseFold, mbnMbcToCode, missing, missing, safeLengthForUptoFour, safeLengthForUptoFourGreatedThan127, safeLengthForUptoThree, safeLengthForUptoTwo, strCodeAt, strLengthasciiApplyAllCaseFold, asciiCaseFoldCodesByString, asciiMbcCaseFold, isCodeCTypeInternalasciiToLower, asciiToUpper, digitVal, equals, getCharset, getIndex, getName, hashCode, isAlnum, isAlpha, isAscii, isAscii, isAsciiCompatible, isBlank, isCntrl, isDigit, isDummy, isFixedWidth, isGraph, isLower, isMbcAscii, isMbcCrnl, isMbcHead, isMbcWord, isNewLine, isPrint, isPunct, isSbWord, isSingleByte, isSpace, isUpper, isWord, isWordGraphPrint, isXDigit, length, load, maxLength, maxLengthDistance, mbcodeStartPosition, minLength, odigitVal, prevCharHead, replicate, rightAdjustCharHead, rightAdjustCharHeadWithPrev, setName, setName, step, stepBack, strByteLengthNull, strLengthNull, strNCmp, toLowerCaseTable, toString, xdigitValstatic final boolean USE_INVALID_CODE_SCHEME
private static final int INVALID_CODE_FE
private static final int INVALID_CODE_FF
public java.lang.String getCharsetName()
getCharsetName in class UnicodeEncodingpublic boolean isNewLine(byte[] bytes,
int p,
int end)
AbstractEncodingisNewLine in class AbstractEncodingpublic int codeToMbcLength(int code)
Encodingcode_to_mbclencodeToMbcLength in class Encodingpublic int mbcToCode(byte[] bytes,
int p,
int end)
Encodingmbc_to_codestatic byte trailS(int code,
int shift)
static byte trail0(int code)
public int codeToMbc(int code,
byte[] bytes,
int p)
Encodingpublic int mbcCaseFold(int flag,
byte[] bytes,
IntHolder pp,
int end,
byte[] fold)
AbstractEncodingmbcCaseFold in class UnicodeEncodingflag - case fold flagpp - an IntHolder that points at character headfold - a buffer where to extract case folded character
Oniguruma equivalent: mbc_case_foldpublic int[] ctypeCodeRange(int ctype,
IntHolder sbOut)
ctypeCodeRange in class Encodingprivate static boolean utf8IsLead(int c)
public int leftAdjustCharHead(byte[] bytes,
int p,
int s,
int end)
leftAdjustCharHead in class Encodingbytes - byte streamp - positions - stopend - endpublic boolean isReverseMatchAllowed(byte[] bytes,
int p,
int end)
isReverseMatchAllowed in class Encoding