Split string with regex based on two conditions(基于两个条件使用正则表达式拆分字符串)-6ren

Split string with regex based on two conditions(基于两个条件使用正则表达式拆分字符串)

转载作者：bug小助手更新时间：2023-10-27 20:53:58

I would like to split this string:

我想把这根弦分开：

lg:[:after]:hover:color-blue

The two conditions are:

这两个条件是：

Split by :

If there's a [], get the content (even if there's a : inside)

The result would be

其结果将是

lg

:after

hover

color-blue

Possible inputs would be:

可能的投入包括：

[:after]:hover:color-blue

hover:color-blue

lg:hover:color-blue

What I have so far:

到目前为止，我所拥有的是：

const regex = /(?:([^\:\[\]]+)|\[([^\[\]]+)\])/g;

const matches = [...string.matchAll(regex)].map((match) =>
  typeof match[2] !== "undefined" ? match[2] : match[1]
);

It works well but the map feels hacky.

它工作得很好，但地图感觉很粗糙。

Is there a way to get the desired output directly from the regex?

有没有办法直接从正则表达式获得所需的输出？

更多回答

Depending on your input you could use split() too, with something like :?\[|\]:?|:(?![^\]\[]*]) though this would leave empty elements if in the input there is a colon/bracket at start or end of the string which would need to be removed.

根据您的输入，您也可以使用Split()，如下所示：？\[|\]：？|：(？！[^\]\[]*])，但如果在输入中字符串的开头或结尾处有冒号/方括号需要删除，则会留下空元素。

that's another approach, thanks

这是另一种方法，谢谢

Is the string enclosed in brackets ([ and ]) always immediately preceded and followed by a colon (:)?

括在方括号([和])中的字符串是否总是前后紧跟冒号(：)？

If they appear before the final :, yes. But I've found a special case, e. g. :hover-[:after]:bg-color-[primary-light]-100... the [primary-light] doesn't need to be matches but the [:after] does.

如果他们出现在决赛之前：是的。但我发现了一个特例，例如：hover-[：after]：bg-color-[primary-light]-100...[主光]不需要匹配，但[：After]需要匹配。

Normally you should not substantially modify your answer after one or more answers have been posted, lest your changes invalidate those answers. In this case, however, there is only one answer and that answer has been updated to address your comment immediately above. Consequently, I suggest you edit your question to include the content of your above comment. Readers should not be required to read through the comments to understand the question.

通常，在发布一个或多个答案后，您不应大量修改您的答案，以免您的更改使这些答案无效。然而，在这种情况下，只有一个答案，该答案已被更新，以解决您的评论紧随其后。因此，我建议您编辑您的问题，以包括您的上述评论的内容。读者不应该被要求通读评论来理解问题。

优秀答案推荐

Yes:

是：

(?<=\[)[^[\]]+(?=\])   # 1+ non-square-brackets inside a pair of those
|                      # or
(?<=^|:)               # a segment
(?:                    # that consists of multiple subsegments,
  \[[^[\]]+\]          # either bracketed
|                      # or
  [^[:\]]+             # non-bracketed,
){2,}                  # 2 or more times.
(?=:|$)                #

Try it on regex101.com.

在regex101.com上试试吧。

This might be shortened as [^[\]\n]+(?=\])|(?:\[[^[\]\n]+\]|[^[:\]\n]+){2,}, which looks less intimidating, if that happens to not match anything unexpected.

这可能会被缩写为[^[\]\n]+(？=\])|(？：\[[^[\]\n]+\]|[^[：\]\n]+){2，}，如果它碰巧与任何意想不到的东西不匹配，看起来就不那么可怕了。

The regex above is, however, prone to backtracking. A "better" version of it would be:

然而，上面的正则表达式很容易回溯。更好的版本应该是：

(?<=\[)[^[\]\n]+(?=\])
|
(?<=^|:)
(?:                    # Before trying to match 2+ subsegments
  (?=([^[:\]\n]+))     # check if the next non-bracketed subsegment
  \1(?=:|$)            # is a full segment, in which case we match
|                      # and skip the 2nd alternative entirely.
  (?:
    \[[^[\]\n]+\]
  |
    [^[:\]\n]+
  ){2,}
)

Try it on regex101.com.

在regex101.com上试试吧。

Since ECMAScript flavor doesn't support atomic groups, we simulated the atomic behaviour using a capturing group inside a lookahead along with the corresponding backreference.

由于ECMAScript风格不支持原子组，因此我们使用前视中的捕获组以及相应的反向引用来模拟原子行为。

Note that we need to use .match() instead of .split():

请注意，我们需要使用.Match()而不是.Split()：

string.match(/[^[\]\n]+(?=\])|(?:\[[^[\]\n]+\]|[^[:\]\n]+){2,}/g)

Try it:

试试看：

console.config({ maximize: true });

const testcases = [
  'lg:[:after]:hover:color-blue',
  '[:after]:hover:color-blue',
  'hover:color-blue',
  'lg:hover:color-blue',
  'hover:[:after]:bg-color-[primary-light]-100',
  'foo:[bar]:[start-but]-not-end',
  'baz-qux:end-[but:not]:start'
];

const regex = /[^[\]\n]+(?=\])|(?:\[[^[\]\n]+\]|[^[:\]\n]+){2,}/g;

for (const testcase of testcases) {
  console.log(testcase, testcase.match(regex));
}

<script src="https://gh-canon.github.io/stack-snippet-console/console.min.js"></script>

Or... you can just split by (?<!\[[^[\]]*): and handle things from there; that would be the "shortest" and clearest solution:

或者..。您只需除以(？<！\[[^[\]]*)：并从那里处理事情；这将是最短的、最清晰的解决方案：

string.split(/(?<!\[[^[\]]*):/).map(
  token => token.match(/^\[.+]$/) ? token.slice(1, -1) : token
)

Try it:

试试看：

console.config({ maximize: true });

const testcases = [
  'lg:[:after]:hover:color-blue',
  '[:after]:hover:color-blue',
  'hover:color-blue',
  'lg:hover:color-blue',
  'hover:[:after]:bg-color-[primary-light]-100',
  'foo:[bar]:[start-but]-not-end',
  'baz-qux:end-[but:not]:start'
];

const regex = /(?<!\[[^[\]]*):/;

for (const testcase of testcases) {
  console.log(
    testcase,
    testcase.split(regex).map(
      token => token.match(/^\[.+]$/) ? token.slice(1, -1) : token
    )
  );
}

<script src="https://gh-canon.github.io/stack-snippet-console/console.min.js"></script>

更多回答

hi, I've found a special case, e. g. : hover:[:after]:bg-color-[primary-light]-100... [primary-light] doesn't need to be matched but [:after] does.

你好，我发现了一个特例，例如：hover：[：after]：bg-color-[primary-light]-100...[主光]不需要匹配，但[：After]需要匹配。

@JoseCarlosRamírez See updated answer.

@JoseCarlosRamírez查看更新的答案。

Thank you for such a thorough answer. After posting the question, I discovered that I need to know if the "modifier" comes from a [] or not, so the split will work better in that case. But it's good to have a solution for the OQ. Thanks!

谢谢你如此透彻的回答。在发布了这个问题后，我发现我需要知道“修饰语”是否来自[]，因此拆分在这种情况下会更好地工作。但有一个OQ的解决方案是很好的。谢谢!

文章推荐： Python Yaml parse inf as float(Python YAML将inf解析为浮点型)

jquery - 正/负最大值输入
我有一个加号/减号按钮，希望用户不能选择超过 20 个但不知道如何让它工作。我尝试使用 min="1"max="5 属性，但它们不起作用。这是我的代码和一个 fiddle 链接。https://jsf
r - ggplot2 正/负图无法清晰呈现
我正在尝试复制顶部底部图，如示例 here但它没有正确渲染(紫色系列有 +ve 和 -ve 值，绿色为负值)留下杂乱的人工制品。我也在努力创建一个玩具示例来复制这个问题，所以我希望尽管我缺乏数据，但有
google-maps - 正/负纬度和经度值与基本方向
已关闭。此问题不符合Stack Overflow guidelines 。目前不接受答案。这个问题似乎与 help center 中定义的范围内的编程无关。 . 已关闭 6 年前。社区去年审查了是
c++在添加两个绝对(正)值时得到负值
这个问题在这里已经有了答案: Adding two positive integers gives negative answer.Why? (4 个答案) 关闭 5 年前。我遇到了一个奇怪的问题
Java将负/正字符串数字转换为负/正 double
有谁知道如何将字符串值类型 -4,5 或 5,4 转换为 double -4.5 或 5.4? 最佳答案只需使用 Double.parseDouble(Locale, String); 糟糕，我很困
python - 在数据框中的新列中返回 TextBlob 正、负或中性分类
我正在尝试根据 TextBlob 分类插入一个仅包含“正”或“负”字符串的新数据框列:对于我的 df 的第一行，结果是 ( pos , 0.75, 0.2499999999999997)我想要' 正
VBA 循环根据相邻单元格更改单元格值(正/负)和字体颜色
我对 VBA 非常陌生，无法理解如何在一个循环中完成 2 个任务。我非常感谢您的帮助。我已经能够根据第 3 列中的数据更改第 2 列中的数值，但我不明白如何将负值的字体更改为红色。表格的大小每月都
jquery - 如何发送 "+"符号(加，正)
欢迎，我正在使用 jquery 通过 POST 发送表单。这就是我获得值(value)的方式。 var mytext = $("#textareaid").val(); var dataStrin
c# - 正 System.Double 值的快速下限和上限替代方案
double d = 0; // random decimal value with it's integral part within the range of Int32 and always p
javascript - 使用 Regex 正/负前瞻替换字符？
我有这个字符串: var a='abc123#xyz123'; 我想构建 2 个正则表达式替换函数: 1) 用 '*' 替换所有确实有 future '#'的字符(不包括'#') 所以结果应该是这样的
android - 如何将 DialogFragment 正/负按钮保留在软键盘上方
我正在使用 DialogFragment。当用户从 Gmail 平板电脑应用程序的屏幕与下面示例图片中的编辑文本进行交互时，我希望正面和负面按钮保持在键盘上方。在我的尝试中不起作用，这是我的 Dia
binary - 二进制补码； 0FFFFh 正，0FFFFh 负？
从组装艺术一书中，我复制了这句话: In the two’s complement system, the H.O. bit of a number is a sign bit. If the H.O
c++ - 基于数值(正、负、零)实现条件表达式的最佳方式
是否有更好更优雅的方法来实现下面的简单代码(diffYear、A 和 B 是数字): diffYear = yearA - yearB; if (diffYear == 0) { A = B
boolean true - 正 1 还是负 1？
我正在设计一种语言，并尝试确定 true 应该是 0x01 还是 0xFF。显然，所有非零值都将转换为 true，但我正在尝试确定确切的内部表示。每种选择的优点和缺点是什么？最佳答案没关系，只要
android - alertdialog 正/负按钮与全屏对话框 fragment 中的父级不匹配
在我的 dialogfragment 类的 OnCreateDialog 中，我正在这样做: AlertDialog.Builder builder = new AlertDialog.Builder
c++ - 正 lambda : '+[]{}' - What sorcery is this?
这个问题在这里已经有了答案: Resolving ambiguous overload on function pointer and std::function for a lambda usin
ios - 正 NSDecimalNumber 返回意外的 64 位整数值
我偶然发现了一个奇怪的 NSDecimalNumber 行为:对于某些值，调用 integerValue、longValue、longLongValue 等，返回意想不到的值(value)。示例: l
c++ - 正 lambda : '+[]{}' - What sorcery is this?
这个问题在这里已经有了答案: Resolving ambiguous overload on function pointer and std::function for a lambda using
regex - 匹配负/正 float/int 正则表达式 - 这是邪恶的吗？
我有这个正则表达式来测试用户输入是否有效: value.length === 0 || value === '-' || (!isNaN(parseFloat(value)) && /^-?\d+\.
matlab - 来自 matlab "fitgmdist"函数的不合理 [正] 对数似然值
我想用高斯混合模型拟合数据集，数据集包含大约 120k 个样本，每个样本有大约 130 个维度。当我使用 matlab 执行此操作时，我运行脚本(簇号为 1000): gm = fitgmdist(d

bug小助手

个人简介

我是一名优秀的程序员,十分优秀！

作者热门文章

滴滴打车优惠券免费领取

全站热门文章

首页

博学

6Ren·AI

商城

Split string with regex based on two conditions(基于两个条件使用正则表达式拆分字符串)