qq_35858961

Explanation of Scoring Metric（for kaggle TGS Salt Identification Challenge）

kaggle竞赛：TGS Salt Identification Challenge中关于评分度量（scoring metric）

原文：https://www.kaggle.com/pestipeti/explanation-of-scoring-metric

评分指标的说明
我发现这场比赛的得分指标的解释有点令人困惑，我想为那些刚刚进入或尚未进入太远的人创建指南。如果你发现错误，请给我留言，谢谢。

Step-By-Step Explanation of Scoring Metric

https://www.kaggle.com/stkbailey/step-by-step-explanation-of-scoring-metric/notebook

Rationale

I found the explanation for the scoring metric on this competition a little confusing, and I wanted to create a guide for those who are just entering or haven't made it too far yet. The metric used for this competition is defined as the mean average precision at different intersection over union (IoU) thresholds.

我发现这场比赛的得分指标的解释有点令人困惑，我想为那些刚刚进入的新手创建指南。用于该竞争的度量被定义为在设置不同交并比（IoU）阈值下的平均精度。

This tells us there are a few different steps to getting the score reported on the leaderboard. For each image...

For each submitted nuclei "prediction", calculate the Intersection of Union metric with each "ground truth" mask in the image.
Calculate whether this mask fits at a range of IoU thresholds.
At each threshold, calculate the precision across all your submitted masks.
Average the precision across thresholds.

Across the dataset...

Calculate the mean of the average precision for each image

这告诉我们在排行榜上报告得分有几个不同的步骤。对于每张图片......

对于每个提交的核“预测”，计算图像中每个“原始图片”掩模的联合度量的交点。
计算此掩码是否适合IoU阈值范围。
在每个阈值处，计算所有提交的蒙版的精度。
平均阈值的精度。
跨数据集......

计算每个图像的平均精度的平均值。

These are the steps to calculate to correct score:以下是计算正确分数的步骤：

For all of the images/predictions对于所有原始图像/预测图片
- Calculate the Intersection of Union metric ("compare" the original mask with your predicted one)
- Calculate whether your predicted mask fits at a range of IoU thresholds.
- At each threshold, "calculate" the precision of your predicted masks.
- Average the precision across thresholds.
Across the dataset
- Calculate the mean of the average precision for each image.

计算联合交点的度量标准（“比较”原始掩码与预测的掩码）
计算您预测的面罩是否适合IoU阈值范围。
在每个阈值处，“计算”预测掩模的精度。
平均阈值的精度。
跨数据集

计算每个图像的平均精度的平均值。

下面我们进行一次模拟预测来深入理解

Picking test image选择测试图片

I'm going to pick a sample image from the training dataset, load the masks, then create a "mock predict"
我将从训练数据集中选择一个样本图像，加载mask，然后创建一个“模拟预测”(只用一张图片)

import matplotlib.pyplot as plt
import numpy as np
import pandas as pd
import imageio

from scipy import ndimage
from pathlib import Path

# Get image
im_id = '2bf5343f03'
im_dir = Path('../input/train/')
im_path = im_dir / 'images' / '{}.png'.format(im_id)
img = imageio.imread(im_path.as_posix())

# Get mask
im_dir = Path('../input/train/')
im_path = im_dir / 'masks' / '{}.png'.format(im_id)
target_mask = imageio.imread(im_path.as_posix())

# Fake prediction mask
pred_mask = ndimage.rotate(target_mask, 10, mode='constant', reshape=False, order=0)
pred_mask = ndimage.binary_dilation(pred_mask, iterations=1)

# Plot the objects
fig, axes = plt.subplots(1,3, figsize=(16,9))
axes[0].imshow(img)
axes[1].imshow(target_mask,cmap='hot')
axes[2].imshow(pred_mask, cmap='hot')

labels = ['Original', '"GroundTruth" Mask', '"Predicted" Mask']
for ind, ax in enumerate(axes):
    ax.set_title(labels[ind], fontsize=18)
    ax.axis('off')

Intersection Over Union (for a single Prediction-GroundTruth comparison)

The IoU of a proposed set of object pixels and a set of true object pixels is calculated as:预测出来的一组对象像素和一组真实对象像素的IoU计算如下：

The intersection and the union of our GT and Pred masks are look like this:
本征图片GT mask和预测图片Pred mask的交集和并集看起来像这样：

代码为：

A = target_mask
B = pred_mask
intersection = np.logical_and(A, B)
union = np.logical_or(A, B)

fig, axes = plt.subplots(1,4, figsize=(16,9))
axes[0].imshow(A, cmap='hot')
axes[0].annotate('npixels = {}'.format(np.sum(A>0)), 
                 xy=(10, 10), color='white', fontsize=16)
axes[1].imshow(B, cmap='hot')
axes[1].annotate('npixels = {}'.format(np.sum(B>0)), 
                 xy=(10, 10), color='white', fontsize=16)

axes[2].imshow(intersection, cmap='hot')
axes[2].annotate('npixels = {}'.format(np.sum(intersection>0)), 
                 xy=(10, 10), color='white', fontsize=16)

axes[3].imshow(union, cmap='hot')
axes[3].annotate('npixels = {}'.format(np.sum(union>0)), 
                 xy=(10, 10), color='white', fontsize=16)

labels = ['GroundTruth', 'Predicted', 'Intersection', 'Union']
for ind, ax in enumerate(axes):
    ax.set_title(labels[ind], fontsize=18)
    ax.axis('off')

So, for this mask the IoU metric is calculated as:

Thresholding the IoU value (for a single GroundTruth-Prediction comparison)

Next, we sweep over a range of IoU thresholds to get a vector for each mask comparison. The threshold values range from 0.5 to 0.95 with a step size of 0.05: (0.5, 0.55, 0.6, 0.65, 0.7, 0.75, 0.8, 0.85, 0.9, 0.95). In other words, at a threshold of 0.5, a predicted object is considered a "hit" if its intersection over union with a ground truth object is greater than 0.5.

阈值化IoU值（对于单个GroundTruth-Prediction比较）这里的评价标准是对应于这个竞赛的
接下来，我们对应于每个mask比对计算出来的IoU值，扫描一系列IoU阈值。阈值范围为0.5至0.95，步长为0.05：（0.5,0.55,0.6,0.65,0.7,0.75,0.8,0.85,0.9,0.95）。换句话说，在阈值为0.5时，如果预测对象与地面实况对象结合的交点大于0.5，则认为该预测对象是“命中”。这样，我们就得到一个对应的vector. 如IoU = 0.6464 这次比对所获得的vector为（1,1,1,0,0,0,0，0,0,0）

Does this IoU hit at each threshold?
0.50     True
0.55     True
0.60     True
0.65    False
0.70    False
0.75    False
0.80    False
0.85    False
0.90    False
0.95    False
Name: GT-P, dtype: bool

这个步骤的代码为：

def get_iou_vector(A, B, n):
    intersection = np.logical_and(A, B)
    union = np.logical_or(A, B)
    iou = np.sum(intersection > 0) / np.sum(union > 0)
    s = pd.Series(name=n)
    for thresh in np.arange(0.5,1,0.05):
        s[thresh] = iou > thresh
    return s

print('Does this IoU hit at each threshold?')
print(get_iou_vector(A, B, 'GT-P'))

Single-threshold precision for a single image单个图像的单阈值精度

At each threshold value t, a precision value is calculated based on the number of true positives (TP), false negatives (FN), and false positives (FP) resulting from comparing the predicted object to all ground truth objects

对于每个阈值t，基于通过将预测对象与标签（mask）对象进行比较而得到的真阳性（TP），假阴性（FN）和误报（FP）的数量来计算精度值。

I think this step causes the confusion. I think the evaluation description of this competition was simply copied from the desc of the 2018 Data Science Bowl (or something similar) competition, where you can predict more than one mask/object per image.

我认为这一步（竞赛度量标准metirc）会引起困惑。我认为本次比赛的评估描述只是简单地从2018年数据科学杯（或类似的）比赛中复制而来，在那里你可以预测每张图像不止一个mask/物体。

According to @William Cukierski's Discussion thread:

https://www.kaggle.com/c/tgs-salt-identification-challenge/discussion/61550#360922

Although this competition metric allows prediction of multiple segmented objects per image, note that we have encoded masks as one object here (see train.csv as an example). You will be best off to predict one mask per image.

Because of we have to predict only one mask per image, the computation below (from the evaluation desc) is a bit confusing:

虽然此竞争指标允许预测每个图像的多个分段对象，但请注意我们在此处将mask编码为一个对象（请参阅train.csv作为示例，即标签区域为1，非标签区域而0，纯黑，相当于二分类）。我们构建模型时每张图像预测一个mask。

因为我们每个图像必须只能预测一个mask，所以下面的计算（来自evaluation desc）有点令人困惑：

This only makes sense if we would have more than one predicted mask/segment per image. In our case I think this would be a better description:

GT mask is empty, your prediction non-empty: (FP) Precision(threshold, image) = 0
GT mask non empty, your prediction empty: (FN) Precision(threshold, image) = 0
GT mask empty, your prediction empty: (TN) Precision(threshold, image) = 1
GT mask non-empty, your prediction non-empty: (TP) Precision(threshold, image) = IoU(GT, pred) > threshold

这只有在每个图像有多个预测的mask/区域（即像宏观的语义分割那样有多个分割）时才有意义。在我们的例子中（二值分割），我认为可以将这个评价标准更好的描述为：

GTmask为空，预测非空：（FP）精度（阈值，图像）= 0 (标签图片为空即为全黑，）
GTmask非空，您的预测为空：（FN）精度（阈值，图像）= 0
GTmask为空，预测为空：（TN）精度（阈值，图像）= 1
GTmask非空，您的预测非空：（TP）精度（阈值，图像）= IoU（GT，pred）>阈值

See @William Cukierski's comment here

If a ground truth is empty and you predict nothing, you get a perfect score for that image. If the ground truth is empty and you predict anything, you get a 0 for that image.

我们重新梳理一下上面的metic : 即如果数据图片(ground truth)为空，即没有salt的图片，我们的预测mask也为空，则，这个也测满分。而不为空时：1,我们的预测为空，则结果错误，0分；2,我们的预测不为空（有mask区域，对应于这个比赛，就是有结果mask不为0的区域），此时，就比对本征图片的标签mask,与预测图片的mask的IoU,>阈值则为1，反之则为0.

Multi-threshold precision for a single image单个图像的多阈值精度

The average precision of a single image is then calculated as the mean of the above precision values at each IoU threshold:对应于上述每个IoU阈值处精度值，求其平均值即单个图像的平均精度值。

Here, we simply take the average of the precision values at each threshold to get our mean precision for the image.
在这里，我们简单地取每个阈值处的精度值的平均值来获得图像的平均精度。

ppt = get_iou_vector(A, B, 'precisions_per_thresholds')
print ("Average precision value for image `{}` is {}".format(im_id, ppt.mean()))

上述代码是选取一张图片计算其精度（例子）

Average precision value for image `2bf5343f03` is 0.3

Let's see what happens if our prediction is 100% accurate.

# `A` is the target mask
ppt = get_iou_vector(A, A, 'GT-P')
print('Does this IoU hit at each threshold?')
print(ppt)
print ("Average precision value for image `{}` is {}".format(im_id, ppt.mean()))

Does this IoU hit at each threshold?
0.50    True
0.55    True
0.60    True
0.65    True
0.70    True
0.75    True
0.80    True
0.85    True
0.90    True
0.95    True
Name: GT-P, dtype: bool
Average precision value for image `2bf5343f03` is 1.0

Mean average precision for the dataset 数据集的平均精度

Lastly, the score returned by the competition metric is the mean taken over the individual average precisions of each image in the test dataset.最后，竞赛度量返回的分数是在测试数据集中每个图像的各个平均精度所取的平均值。

Therefore, the leaderboard metric will simply be the mean of the precisions across all the images.

因此，排行榜度量将仅仅是所有图像的精度的平均值。

第二部分，我们看看竞赛中别人的IoU计算方式。

https://www.kaggle.com/shaojiaxin/u-net-with-simple-resnet-blocks-v2-new-loss

def get_iou_vector(A, B):    #传入ground truth的mask和预测的mask(一次传入一个batch_size大小)
    batch_size = A.shape[0]
    metric = []              #度量列表，列表中每张元素用于存放一个batch_size中每张图片的的度量值
    for batch in range(batch_size):
        t, p = A[batch]>0, B[batch]>0
#         if np.count_nonzero(t) == 0 and np.count_nonzero(p) > 0:
#             metric.append(0)
#             continue
#         if np.count_nonzero(t) >= 1 and np.count_nonzero(p) == 0:
#             metric.append(0)
#             continue
#         if np.count_nonzero(t) == 0 and np.count_nonzero(p) == 0:
#             metric.append(1)
#             continue
        
        intersection = np.logical_and(t, p) #这里用了一个很巧妙的函数来计算交并，
#                 主要是得益于数据的形式，可以查看train.csv文件中的数据，
#               有salt的标签部位有数字，而没有的地方为空（0,），数据只给出了有值的部分。
        union = np.logical_or(t, p)
        iou = (np.sum(intersection > 0) + 1e-10 )/ (np.sum(union > 0) + 1e-10) 
#                 计算一次比对的IoU
        thresholds = np.arange(0.5, 1, 0.05)
        s = []
        for thresh in thresholds:
            s.append(iou > thresh)
#              s是这样的一组值，如前面例子所示IoU值为0.6464那个[1,1,1,0,0,0,0,0,0,0]
        metric.append(np.mean(s))
#              metric实际一次存放一次对比数据的平均值，上面的即为0.3？ 是的
#              最终metirc中保存的是每次对比后的值，大小为一个batch_size
    return np.mean(metric)
#              一个batch_size的 metirc亦是其均值。

def my_iou_metric(label, pred):
    return tf.py_func(get_iou_vector, [label, pred>0.5], tf.float64)

def my_iou_metric_2(label, pred):
    return tf.py_func(get_iou_vector, [label, pred >0], tf.float64)

关于上面这种IoU计算方式，该作者的思路来源于下面：

https://www.kaggle.com/donchuk/fast-implementation-of-scoring-metric

# This Python 3 environment comes with many helpful analytics libraries installed
# It is defined by the kaggle/python docker image: https://github.com/kaggle/docker-python
# For example, here's several helpful packages to load in 

import numpy as np # linear algebra
import pandas as pd # data processing, CSV file I/O (e.g. pd.read_csv)

# Input data files are available in the "../input/" directory.
# For example, running this (by clicking run or pressing Shift+Enter) will list the files in the input directory

import os
print(os.listdir("../input"))

# Any results you write to the current directory are saved as output.

def get_iou_vector(A, B):
    batch_size = A.shape[0]
    metric = []
    for batch in range(batch_size):
        t, p = A[batch], B[batch]
        if np.count_nonzero(t) == 0 and np.count_nonzero(p) > 0:
            metric.append(0)
            continue
        if np.count_nonzero(t) >= 1 and np.count_nonzero(p) == 0:
            metric.append(0)
            continue
        if np.count_nonzero(t) == 0 and np.count_nonzero(p) == 0:
            metric.append(1)
            continue

        intersection = np.logical_and(t, p)
        union = np.logical_or(t, p)
        iou = np.sum(intersection > 0) / np.sum(union > 0)
        thresholds = np.arange(0.5, 1, 0.05)
        s = []
        for thresh in thresholds:
            s.append(iou > thresh)
        metric.append(np.mean(s))

    return np.mean(metric)

如何评估大语言模型生成文本的质量？ gs80140 AI 语言模型人工智能自然语言处理
目录如何评估大语言模型生成文本的质量？1.评估指标概览自动评估指标（AutomaticMetrics）人工评估方法（HumanEvaluation）2.自动评估方法示例（1）计算BLEU分数（2）计算ROUGE分数（3）计算BERTScore（4）使用GPT-4进行评分3.人工评估方法（1）流畅性（Fluency）检查（2）连贯性（Coherence）检查（3）事实准确性（FactualAccur
车辆检测与识别：车辆分类_（9）.车辆分类模型的评估与优化 zhubeibei168 机器人（二）分类数据挖掘人工智能计算机视觉机器学习视频监控
车辆分类模型的评估与优化在车辆检测与识别领域，车辆分类模型的评估与优化是确保模型性能和可靠性的关键步骤。本节将详细介绍如何评估车辆分类模型的性能，并提供一些优化技术，以提高模型的准确性和效率。模型评估指标1.准确率(Accuracy)准确率是最直观的评估指标，表示分类器正确分类的样本占总样本的比例。然而，在不平衡数据集上，准确率可能具有误导性。fromsklearn.metricsimportac
【十自然语言处理项目实战】【10.2 数据收集与预处理】再见孙悟空_ #自然语言处理人工智能知识图谱 transformer 自然语言处理数据收集自然语言处理预处理自然语言处理项目
各位在数据泥潭里打滚的勇士们，今天咱们要聊的这个话题，就像学做川菜必须掌握的"火锅底料炒制法"——数据收集与预处理！这玩意儿看着像脏活累活，实则是决定你模型上限的生死关卡。作为一个曾把BERT训成人工智障的老司机，这就把五年掉坑经验熬成一锅十全大补汤！（戴上橡胶手套准备掏数据）一、数据收集的野路子：比盗墓还刺激的冒险1.1公开数据集寻宝图（附藏宝坐标）①正道的光：Kaggle（数据界的沃尔玛）：搜
常见的数学统计模型若木胡数学模型
以下是常见的数学统计模型分类及简要说明，适用于数据分析、预测和推断等场景：1.参数模型（ParametricModels）假设数据服从特定分布（如正态分布），通过估计参数来描述数据规律。1.1线性回归模型数学形式：(y=\beta_0+\beta_1x_1+\beta_2x_2+\cdots+\beta_px_p+\epsilon)应用：预测连续型目标变量（如房价预测）。特点：简单、可解释性强，假
CAN 调试总结张太行_ arm 网络协议
1.查看CAN设备状态命令：ifconfig~#ifconfigcan0Linkencap:UNSPECHWaddr00-00-00-00-00-00-00-00-00-00-00-00-00-00-00-00UPRUNNINGNOARPMTU:16Metric:1RXpackets:2165errors:0dropped:0overruns:0frame:0TXpackets:0errors:0
[RA-L 2023] Coco-LIC：基于非均匀 B 样条的连续时间紧密耦合 LiDAR-惯性-相机里程计十年一梦实验室 c++
这段代码是一个基于C++的均匀B样条（UniformB-spline）实现，专门用于表示SE(3)变换（即三维空间中的刚体变换，包括旋转和平移）。以下是对代码的总结：1.许可证和版权使用BSD3-ClauseLicense，允许在满足条件的情况下自由分发和修改。版权归VladyslavUsenko和NikolausDemmel所有，属于Basalt项目的一部分。2.功能概述文件定义了一个模板类Se
Prometheus+Grafana监控平台搭建_grafana专业监控项 2401_89828619 prometheus grafana
Prometheus提供多种类型的Exporter用于采集各种不同服务的运行状态。目前支持的有数据库、硬件、消息中间件、存储系统、HTTP服务器、JMX等。·alertmanager警告管理器，用来进行报警。·其他辅助性工具Prometheus系统架构图：它的服务过程是这样的Prometheusdaemon负责定时去目标上抓取metrics(指标)数据，每个抓取目标需要暴露一个http服务的接口给
云原生架构设计理论与实践（14）系统架构
1.云原生背景业务快速发展与开发、运维、运营之间落后的生产关系与生产力的矛盾企业内部各占山头与企业总体战略规划的矛盾企业内部改革，降本增效的需求企业实现数字孪生，数字资产的必然需求企业外部环境，如人工智能发展、安全合规等大环境的要求2.云原生架构的设计原则服务化原则（拆分为微服务、小服务，非功能特性委托）弹性原则（可伸可缩）可观测原则（基于sla，slo，在log，trace，metric三个维度
【深度C++】之“运行时类型识别RTTI” Jinxk8 面向对象C++c++编程语言
0.什么是RTTI运行时类型识别（run-timetypeidentification,RTTI）功能可以获得某类型在运行时的具体动态类型，进而使用该类型的功能。动态类型指的是程序在运行时才可知的类型，与静态类型相对应。静态类型指的是编译时已知的类型。出现静态类型和动态类型定义的原因主要是面向对象的多态。当我们使用父类的指针或引用指向或引用子类对象时，表面上看使用的都是父类的函数，实际上在程序运行
RTTI（Run-Time Type Identification，通过运行时类型识别） Erlei_n c++基础
参考一：RTTI（Run-TimeTypeIdentification，通过运行时类型识别）程序能够使用基类的指针或引用来检查这些指针或引用所指的对象的实际派生类型。RTTI提供了以下两个非常有用的操作符：（1）typeid操作符，返回指针和引用所指的实际类型；（2）dynamic_cast操作符，将基类类型的指针或引用安全地转换为派生类型的指针或引用。面向对象的编程语言，象C++，Java，de
Run-time type information--RTTI diaoju3333 c/c++runtime
Incomputerprogramming,run-timetypeinformationorrun-timetypeidentification(RTTI)[1]referstoaC++mechanismthatexposesinformationaboutanobject'sdatatypeatruntime.Run-timetypeinformationcanapplytosimpledat
ProCmdActionAdd 解读 weixin_39340606 Creo Toolkit c++
ProCmdActionAdd是CreoToolkit中用于定义命令的核心函数，允许开发者向CreoParametric添加自定义操作。这些操作可以与UI元素（如按钮、菜单项或Ribbon命令）关联，当用户与这些元素交互时，绑定的响应函数会被触发，从而实现特定功能。目录参数原型参数详解参数扩展介绍name/action_nameaccess_cb/access_funcaccess_type/pr
【NLP】 5. Word Analogy Task（词类比任务）与 Intrinsic Metric（内在度量） pen-ai NLP 机器学习自然语言处理 word 人工智能
WordAnalogyTask（词类比任务）定义：WordAnalogyTask是用于评估词向量质量的内在指标（IntrinsicMetric）。该任务基于这样的假设：如果词向量能够捕捉单词之间的语义关系，那么这些关系应该能够在向量空间中保持一定的结构。示例：在一个理想的词向量空间中，单词之间的关系应该满足如下等式：king−man+woman≈queenking−man+woman≈queenk
kaggle-ISIC 2024 - 使用 3D-TBP 检测皮肤癌-学习笔记 supernova121 学习笔记
问题描述：通过从3D全身照片(TBP)中裁剪出单个病变来识别经组织学确诊的皮肤癌病例数据集描述：图像+临床文本信息评价指标：pAUC，用于保证敏感性高于指定阈值下的AUC主流方法分析（文本）基于CatBoost、LGBM和XGBoost三者的组合，为每个算法创建了XX个变体，总共XX个模型，进行集成学习。CatBoost在传统梯度提升决策树（GBDT）基础上，引入了一系列关键技术创新，以提升处理类
【模拟面试】计算机考研复试集训（第二天） Albert Edison 计算机考研复试高频考点面试考研职场和发展 c++数据结构算法操作系统
文章目录前言一、专业面试1、OSI参考模型和TCP/IP模型的主要区别是什么？简述各层功能2、什么是瀑布模型？其优缺点是什么？3、什么是递归？使用时需注意什么？4、监督学习与无监督学习的核心区别是什么？请举例说明典型算法5、你在项目中遇到过哪些技术挑战？是如何解决的？二、英文口语1、Canyoutellusaboutatimeyouworkedinateamandfacedchallenges?H
基于python的手写数字识别knn_用sklearn中的KNN实现Kaggle手写数字识别普和司
importcsvfromsklearnimportneighbors#导入训练数据和测试数据defloadData(filename1,filename2,trainDataSet,trainTargetSet,testDataSet):withopen(filename1,'r')ascsvfile1:lines1=csv.reader(csvfile1)dataSet=list(lines1
Linux安装graphite(nginx+uwsgi)过程 caihuan 运维 graphite
由于需要测量程序的各种指标，使用dropwizardmetrics，数据直接输出到graphite.看了很多别人安装graphite的文章，回馈下，写下自己的安装过程。1、查看系统版本cat/proc/versionLinuxversion4.4.10-1-pve(root@elsa)(gccversion4.9.2(Debian4.9.2-10))2、git下载源码Graphite-web:gi
基于线性回归和多项式回归的完整代码 yzx991013 回归线性回归算法
‌1.导入必要库importnumpyasnpimportmatplotlib.pyplotaspltfromsklearn.linear_modelimportLinearRegressionfromsklearn.preprocessingimportPolynomialFeaturesfromsklearn.pipelineimportPipelinefromsklearn.metricsi
kaggle竞赛（初识）薛定谔的码* 人工智能
PART0:Kaggle介绍Kaggle是什么？答案很简单Kaggle是数据挖掘比赛火起来的，以至于中国兴起了很多很多类似的比赛；Kaggle是一个数据科学竞赛的平台，很多公司会发布一些接近真实业务的问题，吸引爱好数据科学的人来一起解决。Kaggle提供了一个介于“完美”与真实之间的过渡，问题的定义基本良好，却夹着或多或少的难点，一般没有完全成熟的解决方案。在参赛过程中与论坛上的其他参赛者互动，能
nodejs部署云服务器数据潜水员 node.js 服务器
###笔记##一、安装Node.js运行环境1.**安装NVM**：```bashbash-c"$(curl-fsSLhttps://gitee.com/RubyMetric/nvm-cn/raw/main/install.sh)"source~/.nvm/nvm.sh```2.**安装Node.js**：```bashnvminstall--lts```3.**检查版本**：```bashnod
51nod 冲刺题 Alaso_shuang OJ选择 OI新手入门刷题 c++算法蓝桥杯
进阶习题：http://www.51nod.com/Question/Index.html#questionId=1481数数字http://www.51nod.com/Challenge/Problem.html#problemId=2151队列复原https://www.51nod.com/Challenge/Problem.html#problemId=3593小明的数字表vector-ST
快速入门：利用fast-elasticsearch-vector-scoring提升ES向量搜索效率劳泉文Luna
快速入门：利用fast-elasticsearch-vector-scoring提升ES向量搜索效率fast-elasticsearch-vector-scoringScoredocumentsusingembedding-vectorsdot-productorcosine-similaritywithESLuceneengine项目地址:https://gitcode.com/gh_mirro
GEE APP——SnowCloudMetrics应用程序，用于分析和可视化雪的覆盖频率（SCF）和雪的消失日期（SDD）此星光明 GEE APP SCF SDD APP 应用雪 gee 交互式界面
目录简介代码解释代码结果网址简介GEEAPP——SnowCloudMetrics应用程序，用于分析和可视化雪的覆盖频率（SCF）和雪的消失日期（SDD）代码解释用于GoogleEarthEngine(GEE)的应用程序，名为"SnowCloudMetrics"，主要用于分析和可视化雪的覆盖频率（SCF）和雪的消失日期（SDD）。以下是其主要功能和结构的具体解释：###主要功能1.**数据导入**：
python3中的os.path模块 hgz_dm 编程语言 python3 os.path
os.path模块主要用于获取文件的属性，这里对该模块中一些常用的函数做些记录。os.abspath(path):获取文件的绝对路径。这里path指的是路径，例如我这里输入“data.csv”[In]os.path.abspath('data.csv')[Out]'E:\\kaggle\\Titanic\\data.csv'os.path.basename(path):获取文件名称。该函数默认通过
基于机器学习的恶意软件检测系统的详细设计与实现源码空间站11 机器学习人工智能课程设计 python 网络安全信息安全恶意软件检测
以下是一个基于机器学习的恶意软件检测系统的详细设计与实现，适合作为课程作业或项目开发。我们将实现一个通过机器学习模型分析恶意软件特征来检测文件是否为恶意软件的系统。总体思路数据准备：选择现有的恶意软件数据集（如Kaggle的恶意软件数据集）或构造模拟数据集。数据集中包含文件的特征（如二进制特征、字符串特征、API调用特征等）和标签（"恶意"或"正常"）。特征提取：提取文件的静态特征（如文件大小、字
OpenTelemetry da__wn 后端
OpenTelemetry简介OpenTelemetry（OTel）是一个开源的可观测性框架，旨在为分布式系统提供标准化的工具和接口，用于生成、收集和管理遥测数据（TelemetryData），包括日志（Logs）、指标（Metrics）和追踪（Traces）。它是CNCF（云原生计算基金会）的孵化项目，融合了早期的OpenTracing和OpenCensus两大标准，成为云原生领域可观测性的事实
langchain4j+ONNX小试牛刀 langchain4j
序本文主要研究一下langchain4j结合ONNX进行得分重排步骤pom.xmldev.langchain4jlangchain4j-onnx-scoring1.0.0-beta1下载模型wgethttps://hf-mirror.com/Xenova/ms-marco-MiniLM-L-6-v2/resolve/main/onnx/model_quantized.onnx?download=t
【K8S问题系列 | 10】在K8S集群怎么查看各个pod占用的资源大小？【已解决】颜淡慕潇 kubernetes 容器云原生后端问题解决
要查看Kubernetes集群中各个Pod占用的资源大小（包括CPU和内存），可以使用以下几种方法：1.使用kubectltop命令kubectltop命令可以快速查看当前Pod的CPU和内存使用情况。需要确保已安装并配置了MetricsServer。查看所有Pod的资源使用情况kubectltoppods--all-namespaces示例输出NAMESPACENAMECPU(cores)MEM
【Elasticsearch】自定义内置的索引生命周期管理（ILM）策略。 risc123456 Elasticsearch elasticsearch
以下是对Elasticsearch官方教程《Customizebuilt-inILMpolicies》的详细解读，结合原文内容，帮助您更好地理解如何自定义内置的索引生命周期管理（ILM）策略。---Elasticsearch教程：自定义内置ILM策略1.背景与目标Elasticsearch提供了内置的索引生命周期管理（ILM）策略，例如`logs@lifecycle`、`metrics@lifec
chatglm3如何进行微调 learner_ctr 人工智能 chatglm3 llm
一、需要的环境内存：因为在loadmodel时，是先放在内存里面，所以内存不能小，最好在30GB左右显存：如果用half()精度来loadmodel的话(int4是不支持微调的)，显存在16GB就可以，比如可以用kaggle的t4gpu，这款性能相当于2070系列，但是显存翻倍python：3.10即可需要安装的包和版本：!pipinstallmodelscope-ihttps://pypi.tu
矩阵求逆（JAVA）初等行变换 qiuwanchi 矩阵求逆（JAVA）
package gaodai.matrix; import gaodai.determinant.DeterminantCalculation; import java.util.ArrayList; import java.util.List; import java.util.Scanner; /** * 矩阵求逆(初等行变换) * @author 邱万迟 *
JDK timer antlove java jdk schedule code timer
1.java.util.Timer.schedule(TimerTask task, long delay)：多长时间（毫秒）后执行任务 2.java.util.Timer.schedule(TimerTask task, Date time)：设定某个时间执行任务 3.java.util.Timer.schedule(TimerTask task, long delay,longperiod
JVM调优总结 -Xms -Xmx -Xmn -Xss coder_xpf jvm 应用服务器
堆大小设置JVM 中最大堆大小有三方面限制：相关操作系统的数据模型（32-bt还是64-bit）限制；系统的可用虚拟内存限制；系统的可用物理内存限制。32位系统下，一般限制在1.5G~2G；64为操作系统对内存无限制。我在Windows Server 2003 系统，3.5G物理内存，JDK5.0下测试，最大可设置为1478m。典型设置： java -Xmx
JDBC连接数据库 Array_06 jdbc
package Util; import java.sql.Connection; import java.sql.DriverManager; import java.sql.ResultSet; import java.sql.SQLException; import java.sql.Statement; public class JDBCUtil { //完
Unsupported major.minor version 51.0（jdk版本错误） oloz java
java.lang.UnsupportedClassVersionError: cn/support/cache/CacheType : Unsupported major.minor version 51.0 (unable to load class cn.support.cache.CacheType) at org.apache.catalina.loader.WebappClassL
用多个线程处理1个List集合 362217990 多线程 thread list 集合
昨天发了一个提问，启动5个线程将一个List中的内容，然后将5个线程的内容拼接起来，由于时间比较急迫，自己就写了一个Demo，希望对菜鸟有参考意义。。 import java.util.ArrayList; import java.util.List; import java.util.concurrent.CountDownLatch; public c
JSP简单访问数据库香水浓 sql mysql jsp
学习使用javaBean，代码很烂，仅为留个脚印 public class DBHelper { private String driverName; private String url; private String user; private String password; private Connection connection; privat
Flex4中使用组件添加柱状图、饼状图等图表 AdyZhang Flex
1.添加一个最简单的柱状图 ? 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 16 17 18 19 20 21 22 23 24 25 26 27 28 <?xml version= "1.0"&n
Android 5.0 - ProgressBar 进度条无法展示到按钮的前面 aijuans android
在低于SDK < 21 的版本中，ProgressBar 可以展示到按钮前面，并且为之在按钮的中间，但是切换到android 5.0后进度条ProgressBar 展示顺序变化了，按钮再前面，ProgressBar 在后面了我的xml配置文件如下： [html] view plain copy <RelativeLa
查询汇总的sql baalwolf sql
select list.listname, list.createtime,listcount from dream_list as list , (select listid,count(listid) as listcount from dream_list_user group by listid order by count(
Linux du命令和df命令区别 BigBird2012 linux
1，两者区别 du，disk usage,是通过搜索文件来计算每个文件的大小然后累加，du能看到的文件只是一些当前存在的，没有被删除的。他计算的大小就是当前他认为存在的所有文件大小的累加和。
AngularJS中的$apply，用还是不用？ bijian1013 JavaScript AngularJS $apply
在AngularJS开发中，何时应该调用$scope.$apply()，何时不应该调用。下面我们透彻地解释这个问题。但是首先，让我们把$apply转换成一种简化的形式。 scope.$apply就像一个懒惰的工人。它需要按照命
[Zookeeper学习笔记十]Zookeeper源代码分析之ClientCnxn数据序列化和反序列化 bit1129 zookeeper
ClientCnxn是Zookeeper客户端和Zookeeper服务器端进行通信和事件通知处理的主要类，它内部包含两个类，1. SendThread 2. EventThread， SendThread负责客户端和服务器端的数据通信，也包括事件信息的传输，EventThread主要在客户端回调注册的Watchers进行通知处理 ClientCnxn构造方法 &
【Java命令一】jmap bit1129 Java命令
jmap命令的用法： [hadoop@hadoop sbin]$ jmap Usage: jmap [option] <pid> (to connect to running process) jmap [option] <executable <core> (to connect to a
Apache 服务器安全防护及实战 ronin47
此文转自IBM. Apache 服务简介 Web 服务器也称为 WWW 服务器或 HTTP 服务器 (HTTP Server)，它是 Internet 上最常见也是使用最频繁的服务器之一，Web 服务器能够为用户提供网页浏览、论坛访问等等服务。由于用户在通过 Web 浏览器访问信息资源的过程中，无须再关心一些技术性的细节，而且界面非常友好，因而 Web 在 Internet 上一推出就得到
unity 3d实例化位置出现布置？ brotherlamp unity教程 unity unity资料 unity视频 unity自学
问：unity 3d实例化位置出现布置？答：实例化的同时就可以指定被实例化的物体的位置,即 position Instantiate (original : Object, position : Vector3, rotation : Quaternion) : Object 这样你不需要再用Transform.Position了, 如果你省略了第二个参数(
《重构，改善现有代码的设计》第八章 Duplicate Observed Data bylijinnan java 重构
import java.awt.Color; import java.awt.Container; import java.awt.FlowLayout; import java.awt.Label; import java.awt.TextField; import java.awt.event.FocusAdapter; import java.awt.event.FocusE
struts2更改struts.xml配置目录 chiangfai struts.xml
struts2默认是读取classes目录下的配置文件，要更改配置文件目录，比如放在WEB-INF下，路径应该写成../struts.xml(非/WEB-INF/struts.xml) web.xml文件修改如下： <filter> <filter-name>struts2</filter-name> <filter-class&g
redis做缓存时的一点优化 chenchao051 redis hadoop pipeline
最近集群上有个job，其中需要短时间内频繁访问缓存，大概7亿多次。我这边的缓存是使用redis来做的，问题就来了。首先，redis中存的是普通kv，没有考虑使用hash等解结构，那么以为着这个job需要访问7亿多次redis，导致效率低，且出现很多redi
mysql导出数据不输出标题行 daizj mysql 数据导出去掉第一行去掉标题
当想使用数据库中的某些数据，想将其导入到文件中，而想去掉第一行的标题是可以加上-N参数如通过下面命令导出数据： mysql -uuserName -ppasswd -hhost -Pport -Ddatabase -e " select * from tableName" > exportResult.txt 结果为： studentid
phpexcel导出excel表简单入门示例 dcj3sjt126com PHP Excel phpexcel
先下载PHPEXCEL类文件，放在class目录下面，然后新建一个index.php文件，内容如下 <?php error_reporting(E_ALL); ini_set('display_errors', TRUE); ini_set('display_startup_errors', TRUE); if (PHP_SAPI == 'cli') die('
爱情格言 dcj3sjt126com 格言
1) I love you not because of who you are, but because of who I am when I am with you. 　　我爱你，不是因为你是一个怎样的人，而是因为我喜欢与你在一起时的感觉。 　　2) No man or woman is worth your tears, and the one who is, won‘t
转 Activity 详解——Activity文档翻译 e200702084 android UI sqlite 配置管理网络应用
activity 展现在用户面前的经常是全屏窗口，你也可以将 activity 作为浮动窗口来使用（使用设置了 windowIsFloating 的主题），或者嵌入到其他的 activity （使用 ActivityGroup ）中。当用户离开 activity 时你可以在 onPause() 进行相应的操作。更重要的是，用户做的任何改变都应该在该点上提交 ( 经常提交到 ContentPro
win7安装MongoDB服务 geeksun mongodb
1. 下载MongoDB的windows版本：mongodb-win32-x86_64-2008plus-ssl-3.0.4.zip，Linux版本也在这里下载，下载地址： http://www.mongodb.org/downloads 2. 解压MongoDB在D:\server\mongodb, 在D:\server\mongodb下创建d
Javascript魔法方法:__defineGetter__,__defineSetter__ hongtoushizi js
转载自： http://www.blackglory.me/javascript-magic-method-definegetter-definesetter/ 在javascript的类中,可以用defineGetter和defineSetter_控制成员变量的Get和Set行为例如,在一个图书类中,我们自动为Book加上书名符号: function Book(name){
错误的日期格式可能导致走nginx proxy cache时不能进行304响应 jinnianshilongnian cache
昨天在整合某些系统的nginx配置时，出现了当使用nginx cache时无法返回304响应的情况，出问题的响应头： Content-Type:text/html; charset=gb2312 Date:Mon, 05 Jan 2015 01:58:05 GMT Expires:Mon , 05 Jan 15 02:03:00 GMT Last-Modified:Mon, 05
数据源架构模式之行数据入口 home198979 PHP 架构行数据入口
注：看不懂的请勿踩，此文章非针对java，java爱好者可直接略过。一、概念行数据入口（Row Data Gateway）：充当数据源中单条记录入口的对象，每行一个实例。二、简单实现行数据入口为了方便理解，还是先简单实现： <?php /** * 行数据入口类 */ class OrderGateway { /*定义元数
Linux各个目录的作用及内容 pda158 linux 脚本
1）根目录“/” 　　根目录位于目录结构的最顶层，用斜线（/）表示，类似于 Windows 操作系统的“C:\“，包含Fedora操作系统中所有的目录和文件。　　2）/bin 　　/bin 　　目录又称为二进制目录，包含了那些供系统管理员和普通用户使用的重要 linux命令的二进制映像。该目录存放的内容包括各种可执行文件，还有某些可执行文件的符号连接。常用的命令有：cp、d
ubuntu12.04上编译openjdk7 ol_beta HotSpot jvm jdk OpenJDK
获取源码从openjdk代码仓库获取(比较慢) 安装mercurial Mercurial是一个版本管理工具。 sudo apt-get install mercurial 将以下内容添加到$HOME/.hgrc文件中，如果没有则自己创建一个： [extensions] forest=/home/lichengwu/hgforest-crew/forest.py fe
将数据库字段转换成设计文档所需的字段 vipbooks 设计模式工作正则表达式
哈哈，出差这么久终于回来了，回家的感觉真好！ PowerDesigner的物理数据库一出来，设计文档中要改的字段就多得不计其数，如果要把PowerDesigner中的字段一个个Copy到设计文档中，那将会是一件非常痛苦的事情。